repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 5d ago
FailproofAI/failproofai
Observability and enforcement for AI agent harnesses. Capture every run and runtime reliability with policy enforcement. 40 built-in policies, a local dashboard, no account required with a generous free cloud plan
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 50%Debugging production agents with Amazon Bedrock AgentCore Observability →
- PossiblePossibly related (embedding) · 49%Reasoning effort, not tool access, buys first-try reliability in agentic code generation: an observational study →
- PossiblePossibly related (embedding) · 49%SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests →
- PossiblePossibly related (embedding) · 49%Falsification, Not Exposure: An Internally Preregistered Placebo-Controlled Decomposition of Self-Repair Feedback in Frozen Small Code Models →
- PossiblePossibly related (embedding) · 57%Failure as a Process: An Anatomy of CLI Coding Agent Trajectories →
- PossiblePossibly related (embedding) · 46%Form, Not Content? A Preregistered, Placebo-Controlled Evaluation of Learned Error-Conditioned Self-Repair Through Prompts and Weights in Frozen Small Code Models →
- PossiblePossibly related (embedding) · 48%Detecting silent agent failures with Amazon Bedrock AgentCore optimization →
Covers
Implements
paperReasoning effort, not tool access, buys first-try reliability in agentic code generation: an observational studypaperSWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction TestspaperFalsification, Not Exposure: An Internally Preregistered Placebo-Controlled Decomposition of Self-Repair Feedback in Frozen Small Code Models
Implements (incoming)
Covers (incoming)
Related across the graph
paperFalsification, Not Exposure: An Internally Preregistered Placebo-Controlled Decomposition of Self-Repair Feedback in Frozen Small Code ModelspaperForm, Not Content? A Preregistered, Placebo-Controlled Evaluation of Learned Error-Conditioned Self-Repair Through Prompts and Weights in Frozen Small Code ModelsnewsDebugging production agents with Amazon Bedrock AgentCore ObservabilitypaperFailure as a Process: An Anatomy of CLI Coding Agent TrajectoriespaperReasoning effort, not tool access, buys first-try reliability in agentic code generation: an observational studypaperSWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction TestsnewsDetecting silent agent failures with Amazon Bedrock AgentCore optimization
