Failure as a Process: An Anatomy of CLI Coding Agent Trajectories
Large language model (LLM) coding agents are increasingly deployed to autonomously perform software engineering tasks in terminal-based environments, making their reliability a growing concern. Existing empirical studies investigate why coding agents fail, yet they largely treat failure as a final outcome rather than a temporal process, providing limited insight into how failures emerge, evolve, and become unrecoverable. We present the first large-scale empirical study of CLI coding-agent failure trajectories, introducing a process-oriented framework that analyzes failure through its onset, ev
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 60%Pluggable by design: An agent mesh for software modernization that adopts the next model release →
- PossiblePossibly related (embedding) · 57%langchain-ai/open-swe →
- PossiblePossibly related (embedding) · 57%dyoshikawa/rulesync →
- PossiblePossibly related (embedding) · 57%FailproofAI/failproofai →
- PossiblePossibly related (embedding) · 54%Revolutionizing software quality: new study explores large language models' pioneering role in defect detection - EurekAlert! →
- PossiblePossibly related (embedding) · 59%When code is abundant →
- LinkedLinked via arxiv author · 85%Xiangxin Zhao →
“Failure as a Process: An Anatomy of CLI Coding Agent Trajectories”
- LinkedLinked via arxiv author · 85%Zihan Liu →
“Failure as a Process: An Anatomy of CLI Coding Agent Trajectories”
