Built at Supabase Select Hackathon 2026 · San Francisco

Your agent will never make the same mistake again.

Every great hire got feedback. Most agents get abandoned instead.

01Simulate
02Observe
03Flag
04Improve
05Test & ship
AGENTS WANT TO BE USED

Everyone is fighting over the 15%.

When Ajay Banga became CEO of Mastercard, he noticed the company's slogan called it the heart of commerce. Yet inside the building, everyone talked about Visa and American Express, rivals competing for the small share of payments that were already electronic.

Almost no one talked about the real competitor: cash, which at the time handled more than 85% of consumer transactions worldwide.¹

So Banga rewrote the mission in two words:kill cash.

Mastercard stopped fighting over the 15% and went after the 85%.

Burning cashFIG. 01 — THE REAL COMPETITOR WAS NEVER VISA.
THE AGENT MARKET, TODAY

AI agents are in the same spot today.

In the US, the most AI-native market on earth, 64% of adults use AI, a number that barely moved from 61% a year earlier.² Only about 24% of AI users use an agent regularly.³ That means roughly 15% of American adults rely on AI agents.⁴ Worldwide, consumer AI reaches about 24% of the population, so roughly 3 in 4 people don't use AI at all.²

The industry is building evals, frameworks and dashboards for the 15% who already use agents. The bigger market is everyone who tried once, got stuck, and never came back.

15%already use agents
85%tried once, got stuck, never came back
← NOBODY IS BUILDING FOR THEM
Evals · frameworks · dashboardsThe drop-off: people who tried an agent once and never came back
OUR MISSIONKill the drop-off.
SOURCES

¹ Carolyn Dewar, Scott Keller and Vikram Malhotra, CEO Excellence: The Six Mindsets That Distinguish the Best Leaders from the Rest (Scribner, 2022). Figures as reported in the book for that historical period.

² Menlo Ventures, 2026: The State of Consumer AI, survey of 5,067 US adults conducted with Morning Consult, July 2026. menlovc.com/perspective/2026-the-state-of-consumer-ai

³ Same Menlo Ventures report, as compiled by Digital Applied, "What People Let AI Agents Access: 2026 Survey Numbers" (Sept 2026).

⁴ itera.ai estimate: 64% of US adults using AI × 24% of AI users using an agent regularly ≈ 15%.

THE PROBLEM

Users don't quit because AI is bad. They quit because nobody fixed it.

Every team shipping AI agents runs into the same pattern:

01 · THE AGENT
Can we move my hearing to Friday?
Our office hours are 9am–6pm.

Misreads one message.

02 · THE USER

Gives up.

03 · THE MANAGER

Notices, but can't edit a prompt.

04 · THE ENGINEER

Can fix the prompt, but never sees the conversation.

Weeks go by, and the same mistake hits the next hundred users.WEEK 1 → WEEK 6 · ×100 USERS
HOW IT WORKS

One loop, from bad reply to better agent.

1
Simulate

Chat with your agent (receptionist, sales, support) right inside itera. Every conversation is real data.

2
Observe

Every turn becomes a trace: input, output, prompt version, model, latency. Click any log to open the full conversation.

3
Flag

Select any message, mark it 👍 or 👎, and explain what should have happened, in plain language. No prompt engineering required.

4
Improve

Feedback goes into a self-improvement queue. itera consolidates it and runs a prompt optimizer (GEPA-style) against the exact version that served that user.

5
Test and ship

The new version runs a battery of synthetic conversations. You review the diff, see why each change was made, and promote it from staging to production.

Back to step 1. Every conversation teaches the agent.
Receptionist agentprompt v12 · production
Simulating
TODAY · SIMULATED SESSION #1,284

Hi, this is Almeida & Costa Law. How can I help you today?

Hi! Can we move my hearing prep call to Friday?

Our office is open Monday to Friday, 9am to 6pm. Is there anything else I can help with?

trace_8f2a · gpt-4.1 · v12 · 1.4s
What should have happened?
MMarina · Office manager

She asked to reschedule, not for our hours. The agent should check the calendar and offer the open Friday slots.

Plain language. No prompt engineering.Send to improvement queue
Improvement queue: 7 feedbacks → consolidating → optimizing v12 → v13 (staging)
FEATURES

No black-box rewrites.

Humans stay in charge of what ships. itera does the prompt engineering.

Prompt diffs with reasons.

Every changed line links back to the feedback that caused it. No black-box rewrites.

receptionist.promptv12 → v13 +3 −1
12 You are the receptionist for Almeida & Costa Law.
13-When users ask about scheduling, share office hours.
14+When users ask to reschedule, check the calendar#214 · Marina
15+and offer the next three open slots.
16 Always confirm the client's name and case number.
17+If the request is unclear, ask one clarifying question.#219 #223 · 3 similar
Why: 3 feedbacks consolidated · optimized against v12 · 48/50 synthetic tests passed

Staging for prompts.

Nothing reaches users until a human approves it.

STAGINGprompt v13
Awaiting review
human approval required
PRODUCTIONprompt v12
Live
Review diffPromote

Synthetic test battery.

See how the new version handles real scenarios before it goes live.

48/50v13 vs v12: +9 passed
Reschedule a hearing
New client intake
Angry client, wrong case no.
Asks for fees in Spanish

Full trace history.

Every input and output, from both AI and humans, versioned and searchable.

reschedule version:v12
14:02:11AIv12Our office is open Monday…
14:02:40HUMAN—👎 She asked to reschedule…
14:05:03AIv13I can move it. Friday 10am…
14:05:20USER—Friday 10am works, thanks!

Built for the people who see the problem.

Managers and reviewers give feedback. itera does the prompt engineering.

M
MARINA · MANAGER"It should offer Friday slots."
R
RAFAEL · REVIEWER"Too formal for WhatsApp."
itera rewrites the prompt → v13
BUILT WITH
Supabase
·
Vercel AI SDK

Agents don't need better benchmarks.
They need users who stay.

Every conversation teaches your agent something. itera makes sure it learns.

Start iteratingBuilt at Supabase Select Hackathon 2026 · San Francisco