repoGitHubTrust 82 · PrimaryPublished 11h agoLive · 11h ago
harleyszhang/rapid_llm
A light llama-like llm inference framework based on the triton kernel.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 61%Show HN: Goku – WASM (wllama)-powered LLM inference and model manager →
- PossiblePossibly related (embedding) · 60%WebLLM: high-performance in-browser LLM inference engine →
- PossiblePossibly related (embedding) · 60%The efficient frontier of LLM inference →
- PossiblePossibly related (embedding) · 59%Hetzner is working on LLM Inference →
- PossiblePossibly related (embedding) · 57%OpenAI and Broadcom unveil LLM-optimized inference chip →
