Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago

DaoyuanLi2816/llm-gpu-lab

One GPU. Full LLM workflow. Real benchmarks. No cloud required.

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

Covers

Implements

Covers (incoming)

Related across the graph

newsAI Model Co-Design: Hardware-Friendly LLM Design | NVIDIA Technical Blog - NVIDIA DevelopernewsBonsai 27B: 1-bit dense LLM running locally in your browser using custom WebGPU kernelsnewsUltra budget 20GB vram with 448GB/s for $100 bucks.newsShow HN: PantheonGPU – GPU health testing and AI workload benchmarkingnewsNASA Puts Google’s Gemma Large Language Model in OrbitnewsAI-Assisted GPU Porting of a 250k Line Legacy Weather Simulation Codenews[Paper] Automated Tensor Scheduling for Hybrid CPU-GPU LLM Inference on Consumer DevicesnewsIf you had a bunch of GPUs lying around, what would you actually build with them? (Running LLMs is off the table) [D]paperWattGPU: Predicting Inference Power and Latency on Unseen GPUs and LLMsnewsKicking off GPU Mode [D]newsBest Local LLMs You Can Run on a Single 24GB GPU in 2026: Qwen, Gemma, Mistral, DeepSeek Compared - MarkTechPostnewsWe'll benchmark an Open weights LLM on any GPU you choose — drop your model + hardware and we'll run it. [D]newsI have a mid-sized GPU cluster and was thinking about giving free compute [D]newsMSI Pro Max Edge AI+ Mini PC Runs 120B Local AI Models With 128GB RAM - HotHardwarenewsRun Massive-Scale UMAP in Minutes Using Multiple GPUs—Without Losing Accuracy | NVIDIA Technical Blog - NVIDIA DevelopernewsShow HN: Computable – Buy, sell, and redeem GPU for the exact weeks you wantnewsMeasuring PCIe transfer under dual GPU with pipeline & tensor llama.cppnewsNVIDIA offers to help Vietnam build own LLM, expand AI GPU capacity - TNGlobalnews100$ worth of gpu runs qwen 3.8 27b at 7.39 t/snewsLinux Kernel Graphics Driver Proposed For GlandaGPU: An Open-Source Soft GPU Core - PhoronixnewsTop Cost-Effective Enterprise GPU Cloud Platforms for AI Workloads with H100–GB200, Elastic Scaling and Pay-as-You-Go Compute - Scott Coop

Topics