Context Theory Get your growth audit

Answer

How do you recover when an AI agent goes down the wrong path?

Stop it, revert to the last known-good state, and restart with a corrected brief. Arguing it back on course rarely works.

Stop, revert to the last good state, and start again with a brief that names the misunderstanding. Correcting an agent mid-run leaves the wrong reasoning present, and it keeps influencing what happens next.

The instinct when a run goes wrong is to correct it: point out the error, ask it to fix the approach, continue. This works for a small deviation and fails for a structural one, and the reason is mechanical rather than psychological. Everything produced so far remains present, including the reasoning that led to the wrong path and the artefacts built on it. Subsequent work is done in the presence of all of it, and the correction is one instruction competing against an accumulated body of contrary material.

The reliable move is therefore to restart rather than to redirect, and the thing that makes restarting cheap is having somewhere to restart from. This is where the working practices pay off: a branch, a snapshot, a copy, a state file written at the end of the last good stage. Teams that have these treat a wrong path as a twenty-minute loss. Teams that do not treat it as a decision between accepting compromised work and redoing several hours of it, and they usually accept the compromised work.

Before restarting, extract what was learned. The failed run almost always discovered something real — a constraint nobody knew about, a file that is not where it should be, a case the brief did not cover. Writing that down before discarding the session is what makes the second attempt better rather than merely different. Discarding without extracting is how teams end up running the same wrong path twice.

The corrected brief should name the misunderstanding explicitly rather than adding general caution. If the agent assumed the data was clean, say it is not and say what is wrong with it. If it optimised for the wrong thing, say which thing. A brief that only becomes more emphatic is addressing effort, which was never the problem, while the specific false assumption remains available to be made again.

There is a judgement call about how early to stop, and the error is nearly always stopping too late. The cost of a wrong path grows with everything built on top of it, so the expected cost of stopping a run that turns out to have been fine is small and the expected cost of letting a wrong one continue is not. In practice the signal is usually available early: the agent's description of what it is doing stops matching what you asked for, and that divergence is visible in the first stage output if anyone looks at it.

Finally, treat a wrong path as information about the setup rather than about the run. One agent misreading one brief is noise. The same misreading twice is a defect in the brief, the material or the tools, and it will keep happening until something outside the agent changes. The most useful thing to do after a recovery is to ask which of those three would have prevented it.

You cannot talk an agent out of a misunderstanding it has already built on, because the building is now part of what it is reasoning from.

Siddharth Sharma, Context Theory

Related questions

Is it ever right to correct in place?

For a local mistake, yes: a wrong value, a misread instruction on one item, a formatting error. The distinction is whether subsequent work depends on the error. If it does not, correcting is cheap and fine. If it does, the correction has to propagate through everything built since, and that propagation is exactly what does not happen reliably.

How do you avoid the same wrong path on the retry?

Name it in the brief as a prohibition tied to a reason, not as an instruction to be careful. The failed run gave you the specific false assumption, which is a rare and valuable thing to have, and stating it directly is far more effective than any general framing. Most second attempts fail for the same reason as the first because the brief was strengthened rather than corrected.

METHOD

Every figure below carries its source and the date it was verified. Nothing on this page is asserted.

The numbers on this page.

Datapoints
What Value Specific to
Sub-15-minute compliance — automated routing vs manual only62.5% vs 39.1%Category-wide
Agents who give up after one contact44%Category-wide

2026 speed-to-lead benchmark · verified

Multi-study aggregate · verified

What is specific to this page.

Evidence
Kind Claim Check it against
WorkflowCorrecting an agent mid-run leaves the reasoning that produced the wrong path present alongside the artefacts built on it, so the correction competes with an accumulated body of contrary material rather than replacing it.Comparing outcomes when a structurally wrong run is corrected in place against when it is restarted from a clean state with the same correction.
ProcurementThe affordability of restarting is determined by whether a known-good state exists to restart from, which converts a wrong path from a choice between compromised work and hours of rework into a bounded loss.Timing recovery from an aborted run with a branch or snapshot available against recovery without one.
ResponseA failed run typically discovers a real constraint or uncovered case, so extracting what was learned before discarding the session is what distinguishes a better second attempt from a merely different one.Comparing second attempts made with and without a written record of what the failed run found.
SoftwareA repeated misreading indicates a defect in the brief, the material or the tools rather than variance in the run, and will recur until one of those three changes, which makes the post-recovery question a design question rather than a retry decision.Recording the specific misunderstanding on each failed run and checking whether it recurs across different sessions.

Each row would be wrong on another industry's page. Where a sourced figure exists it is in the table above instead; these are the constraints that shape the work and do not happen to be numbers.

Start with the measurement.

Reading about a benchmark is not the same as knowing your own number. The audit produces yours, measured rather than estimated.

Get your growth audit

$497 · delivered in 5 business days · credited against month one