repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago
Andyyyy64/whichllm
Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 53%Evaluate a model properly →
- PossiblePossibly related (embedding) · 49%PACE: A Proxy for Agentic Capability Evaluation →
- PossiblePossibly related (embedding) · 49%OpenAI and Broadcom announce chip designed for LLM inference at scale →
- PossiblePossibly related (embedding) · 47%Best tps can I get with Qwen3.5 122B on 32GB VRAM + 64GB RAM? →
- PossiblePossibly related (embedding) · 47%I mapped which local LLMs actually fit each RAM tier, 8 to 128GB (open dataset) →
- PossiblePossibly related (embedding) · 49%Why goodput matters more than throughput for LLM serving →
Related to
Implements
Covers
Covers (incoming)
Related across the graph
newsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsI mapped which local LLMs actually fit each RAM tier, 8 to 128GB (open dataset)paperPACE: A Proxy for Agentic Capability EvaluationnewsBest tps can I get with Qwen3.5 122B on 32GB VRAM + 64GB RAM?tutorialEvaluate a model properlynewsWhy goodput matters more than throughput for LLM serving
