
Failure as a Process: An Anatomy of CLI Coding Agent Trajectories
Rather than just measuring whether coding agents succeed or fail, this large-scale study examines how failures unfold over time across nearly 1,800 annotated agent trajectories. The process-oriented view reveals that many failures aren't sudden — they have identifiable onset points, predictable escalation patterns, and windows where recovery is still possible. Essential reading if you're building or operating coding agents and want to understand where interventions would actually help.
Takeaways3
- Agent failures are temporal processes with identifiable early warning patterns, not just binary outcomes.
- Many failure trajectories have recovery windows that current agents consistently miss, suggesting intervention points for scaffolding improvements.
- Different frontier models fail in structurally distinct ways, meaning model choice affects failure mode, not just success rate.










