time
Seconds on answers you already know
Every repeat waits on model latency, tool calls and orchestration.
Reusable computation for AI agents
Wont finds the work your AI agents keep repeating and turns it into a verified, instant path. When it isn’t sure, it hands the task back to your agent.
Self-hosted · Python SDK · your agent stays in control
01The hidden tax on agent systems
Every run pays full price in model latency, tokens and orchestration, even when your agent has solved the same task many times.
time
Every repeat waits on model latency, tool calls and orchestration.
tokens
The work has the same shape every time. The bill arrives every run.
consistency
Re-deriving a result each time invites drift where you want none.
02Measured, not promised
agent 1,365 ms
1.6 ms
Typical response time (p50)
agent 30 calls
0
Model calls on the compiled path
agent 30/30
30/30
Correct, with no accuracy traded for speed
naive matcher 9 wrong
0 wrong
On 35 tricky lookalike cases
Controlled benchmark · one supported workflow · 30 measured runs per lane · 35 adversarial cases for the lookalike test
03How it works
Wont records what your agent does. Prompts, responses and secrets are stripped out first, so only the safe shape of the work is stored.
It spots the same work recurring with new inputs. Each repeat adds evidence; one sighting is never enough.
The stable part becomes a fast, deterministic path: the groove that repetition wears in.
The path is checked against your agent’s own results before it is ever trusted to run.
Familiar, verified work takes the compiled path in milliseconds. Anything new or uncertain goes to your agent.
Same steps, new inputs: compiled once, verified, reused.Fresh reasoning stays with your agent.
Not a cache. Not a router. Wont reuses the computation, not the answer, so new inputs still get correct results.
04The Reusable Computation Map
Recurring workflows, how often they repeat, which are ready to compile, and every time Wont chose to step aside.
recurring workflows
12runs that repeat
41%compiled paths
3fallbacks this week
7model calls avoided
1,284| workflow | repeats | status | verification |
|---|---|---|---|
| refund_status_lookup | 318 | compiled | 30/30 agree |
| order_total_calc | 204 | compiled | 30/30 agree |
| invoice_summary | 57 | observing | collecting evidence |
| ticket_triage | 23 | stays with agent | novel reasoning |
Preview with illustrative data. The hosted dashboard ships with the public beta; today the same findings come from the SDK’s repetition report.
05What happens when Wont is wrong?
It never blindly replaces your agent. When anything is uncertain, the task goes to your agent: a slower answer, never a silent wrong one.
A path is checked against your agent’s results before it can run.
Shadow mode runs Wont alongside your agent and only compares. Early access.
Unfamiliar or changed inputs go straight to your agent.
No writes to production systems.
measured
A naive matcher executed 9 of them wrongly.
06Built for developers
Run Wont next to your stack, send it your agent traces, and get a report of the work worth compiling.
The SDK ships as the reflex package today. The wont name is coming.
# 1 · install
pip install -e .
# 2 · database
docker compose up -d postgres
export DATABASE_URL="postgresql+psycopg://reflex:reflex_dev_only@localhost:5432/reflex"
python -m alembic upgrade head
# 3 · project + one-time API key
python -m examples.week7_bootstrap_project \
--workspace demo --workspace-name "Demo" \
--project support --project-name "Support" \
--environment development
# 4 · API
python -m uvicorn apps.api.main:app --port 8000
import os
from reflex import Reflex
client = Reflex(
api_key=os.environ["REFLEX_API_KEY"],
base_url=os.environ["REFLEX_BASE_URL"],
)
# trace: a TraceIngestRequest from your agent run
client.traces.ingest(trace, workflow_version="support-v1")
report = client.repetition_report(limit=10)
print(report.repeated_family_count, report.repetition_coverage)
# the quickstart's six demo traces
scoped_execution_count 6
distinct_procedure_count 3
repeated_family_count 2
repeated_execution_count 5
repetition_coverage 5/6
top_families 3, 2
07Where it fits
Status checks, record fetches, routine reads.
Pricing, conversions and totals the model re-derives every time.
Multi-step procedures that keep the same shape.
Supported today: read-only lookups and deterministic calculations. More patterns in design-partner pilots.
08Questions
No. Your agent handles all new work and anything uncertain. Wont only takes over work it has already verified.
A safe trace structure. Raw prompts, responses and credential-like values are stripped before anything is stored.
Changed inputs or shapes fail Wont’s checks, so the task goes back to your agent automatically.
Not today. Wont is read-only by design.
Any agent that can send traces through the SDK or REST API. Wont doesn’t depend on a specific model.
Yes. Start with the repetition report. It only reads the traces you send and never touches how your agent runs.
wont /wōnt/ noun · one’s customary behaviour; habit.