MEND AI · MODEL v1.0.0
Watch Mend AI learn.
One verified fix at a time.
Mend AI is the small model inside every Mend server. When a screen in your app gets stuck, it chooses one of the recoveries you approved, applies it, and learns from whether the browser confirms the next screen really appeared.
When it is not sure, it asks Claude. Claude’s checked answer can only make Mend AI more careful, never take over. Confidence comes from browser-verified results alone.
LIVE LEARNING FEED
Live activity starts with the next server release on October 11, 2026.
From then on, this page will show Mend AI learning from real, browser-verified fixes as they happen: which kinds of failures it learned from, which recovery worked, how many fixes it now makes on its own, and how many corrections it took from Claude.
- Anonymized counts by category only: never names, labels, URLs, element IDs or which customer it came from.
- Small categories are hidden so no single customer’s activity can be picked out.
- Any workspace can opt out, and turning learning off in a workspace stops it learning there.
THE LOOP
Train → fix → iterate.
- 1 · TrainLearns from verified fixes
Every recovery the browser confirms (or rejects) updates what Mend AI knows about that kind of failure, on that server.
- 2 · FixClaude corrects, cautiously
When Mend AI is unsure or a recovery fails, it asks Claude. Only the action Claude chose is counted, never Claude’s text, and it can only make Mend AI more careful.
- 3 · IterateImproves every night
A nightly job tests Mend AI on synthetic transitions and proposes better rules. Each change is reviewed before release and becomes a new version.
BENCHMARK · v1.0.0
Where v1 starts.
Measured on Mend’s synthetic test suite of transition failures, not on customer data. These are the numbers each new version has to beat.
- Fixed and verified
- 63.9%
- Asked Claude
- 63.4
- Wrong recoveries
- 16.1%
of synthetic failures recovered with a browser-verified result.
times per 100 transitions. The goal is to bring this down as Mend AI learns.
tried a recovery that did not verify. Mend reports these to the developer instead of hiding them.
VERSION HISTORY
Mend AI versions.
- v1.0.0October 10, 2026
First release. A learned micro model: outcome counts and safe rules that choose among developer-approved recoveries, with Claude for the cases it is unsure about.
- v1.xNext
Nightly train → fix → iterate improvements, each reviewed and released as a new version with its benchmark results listed here.
ROADMAP
From micro model to flagship local LLM.
- Nowv1 · Micro model
Runs on every Mend server. Learns from verified fixes. Asks Claude only when unsure.
- Nextv1.x · Sharper micro model
More fixes handled locally, fewer Claude calls, and faster recoveries as the nightly loop and real verified outcomes improve it.
- Plannedv2 · Flagship local LLM
A small language model that runs on your own server or device, trained on browser-verified outcomes and developers’ own registered fixes — never on Claude’s text. Most fixes would then need no cloud AI at all.
Honest status: today’s Mend AI v1 is a learned micro model, not yet a language model. v2 is planned and not built yet.