repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 29d ago
notwitcheer/llm-bench-rig
Dual-engine (llama.cpp + vLLM) LLM benchmarking pipeline for GGUF & safetensors on NVIDIA GPUs — speed, quality, live dashboard, publishable cards.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 54%One-Step Gradient Delay is Not a Barrier for Large-Scale Asynchronous Pipeline Parallel LLM Pretraining →
- PossiblePossibly related (embedding) · 51%Bolt Graphics GPU will have 2 DDR5 laptop DIMM slots →
- PossiblePossibly related (embedding) · 51%Nous Research's NousCoder-14B is an open-source coding model landing right in the Claude Code moment →
- PossiblePossibly related (embedding) · 50%Going from single GPU to dual GPU is nice but not in the way I expected →
- PossiblePossibly related (embedding) · 50%Are there good closed vs open LLM rankings? Also, are 70B–350B models actually worth it? →
- PossiblePossibly related (embedding) · 58%We'll benchmark an Open weights LLM on any GPU you choose — drop your model + hardware and we'll run it. [D] →
- PossiblePossibly related (embedding) · 53%Anthropic’s Claude Available in Microsoft Corporation (MSFT) Foundry Powered by Nvidia GPUs - Insider Monkey →
- PossiblePossibly related (embedding) · 46%Diffusion LLM Learns to Be Its Own Draft Model: NVIDIA Releases Tri-Mode Open Weights - Tech Times →
Implements
Covers
newsBolt Graphics GPU will have 2 DDR5 laptop DIMM slotsnewsNous Research's NousCoder-14B is an open-source coding model landing right in the Claude Code momentnewsGoing from single GPU to dual GPU is nice but not in the way I expectednewsAre there good closed vs open LLM rankings? Also, are 70B–350B models actually worth it?
Covers (incoming)
newsWe'll benchmark an Open weights LLM on any GPU you choose — drop your model + hardware and we'll run it. [D]newsAnthropic’s Claude Available in Microsoft Corporation (MSFT) Foundry Powered by Nvidia GPUs - Insider MonkeynewsDiffusion LLM Learns to Be Its Own Draft Model: NVIDIA Releases Tri-Mode Open Weights - Tech TimesnewsAI Model Co-Design: Hardware-Friendly LLM Design | NVIDIA Technical Blog - NVIDIA DevelopernewsHy3 (295B MoE) and NVIDIA Nemotron-Labs-Audex-30B-A3B (audio-capable 30B MoE) GGUF quantsnewsFBI considers deploying AI LLM supercomputers with Nvidia B300 GPUs or Google TPUs - Data Center DynamicsnewsAMD and Cerebras join forces against Nvidia’s Groq LPUsnewsIf you had a bunch of GPUs lying around, what would you actually build with them? (Running LLMs is off the table) [D]
Related across the graph
newsAI Model Co-Design: Hardware-Friendly LLM Design | NVIDIA Technical Blog - NVIDIA DevelopernewsDiffusion LLM Learns to Be Its Own Draft Model: NVIDIA Releases Tri-Mode Open Weights - Tech TimesnewsHy3 (295B MoE) and NVIDIA Nemotron-Labs-Audex-30B-A3B (audio-capable 30B MoE) GGUF quantsnewsNous Research's NousCoder-14B is an open-source coding model landing right in the Claude Code momentnewsIf you had a bunch of GPUs lying around, what would you actually build with them? (Running LLMs is off the table) [D]newsAnthropic’s Claude Available in Microsoft Corporation (MSFT) Foundry Powered by Nvidia GPUs - Insider MonkeynewsFBI considers deploying AI LLM supercomputers with Nvidia B300 GPUs or Google TPUs - Data Center DynamicsnewsWe'll benchmark an Open weights LLM on any GPU you choose — drop your model + hardware and we'll run it. [D]newsAre there good closed vs open LLM rankings? Also, are 70B–350B models actually worth it?newsAMD and Cerebras join forces against Nvidia’s Groq LPUspaperOne-Step Gradient Delay is Not a Barrier for Large-Scale Asynchronous Pipeline Parallel LLM PretrainingnewsBolt Graphics GPU will have 2 DDR5 laptop DIMM slotsnewsGoing from single GPU to dual GPU is nice but not in the way I expected
