SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests
Large language model (LLM)-based software engineering agents are increasingly developed to resolve software issues by generating patches from issue reports and code repositories. Bug reproduction tests (BRTs) are an important building block for such agents and have been shown useful for patch validation. However, it remains unclear whether BRTs can also help the more central stage of patch generation. We first conduct a preliminary study and find that directly using advanced BRT generators to guide patch generation is not beneficial: fail-to-fail BRTs can mislead agents, while even fail-to-pas
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- LinkedLinked via unknownagent-tools →
- LinkedLinked via unknownPatch the Planet: a Daybreak initiative to support open source maintainers →
- LinkedLinked via unknownDebugging production agents with Amazon Bedrock AgentCore Observability →
- LinkedLinked via unknownAgentTrace →
- PossiblePossibly related (embedding) · 48%bug-ops/zeph →
- PossiblePossibly related (embedding) · 55%mozilla/bugbug →
- PossiblePossibly related (embedding) · 46%promptfoo/promptfoo →
