repoGitHubTrust 82 · PrimaryPublished 29d agoLive · 21d ago
hermes-labs-ai/little-canary
little-canary is a prompt-injection detector that reads attacks by their effect on a sacrificial canary model before they reach production. Puts a small canary model in front of your app, watches whether untrusted input compromises it, and returns block, flag, or pass as an inbound preflight check before your primary model acts.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 45%Hermes Agent Security: 7-Layer Defense Setup Guide - Hostinger →
