repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago
Vishisht16/Humane-Proxy
Lightweight AI safety middleware that protects humans by intercepting self-harm and criminal intent in LLM prompts. Features a 3-stage safety pipeline, MCP server for agents, and automated care responses.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 52%Behind the Refusal: Determining Guardrail Activation via Behavioral Monitoring →
- PossiblePossibly related (embedding) · 52%Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI →
- PossiblePossibly related (embedding) · 49%Verisight →
- PossiblePossibly related (embedding) · 48%Orca provides safety layer for autonomous AI agents - Let's Data Science →
- PossiblePossibly related (embedding) · 48%AI browsers can be lulled into a dream world where guardrails no longer apply →
- PossiblePossibly related (embedding) · 50%When Local Monitors Miss Compositional Harm: Diagnosing Distributed Backdoors in Multi-Agent Systems →
Implements
Covers
Related to
Implements (incoming)
Related across the graph
newsOrca provides safety layer for autonomous AI agents - Let's Data SciencepaperWhen Local Monitors Miss Compositional Harm: Diagnosing Distributed Backdoors in Multi-Agent SystemspaperBehind the Refusal: Determining Guardrail Activation via Behavioral MonitoringnewsNemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AInewsAI browsers can be lulled into a dream world where guardrails no longer applycompanyVerisight
