repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 29d ago
john-rocky/apple-silicon-llm-bench
Neutral, reproducible benchmark for local LLMs on Apple Silicon (Mac · iPhone · iPad) — MLX, llama.cpp, CoreML, Apple Foundation Models
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 50%Ask HN: MacBook vs. Dedicated GPU for LLM →
- PossiblePossibly related (embedding) · 64%OpenAI and Broadcom announce chip designed for LLM inference at scale →
- PossiblePossibly related (embedding) · 59%Evaluate a model properly →
- PossiblePossibly related (embedding) · 48%I mapped which local LLMs actually fit each RAM tier, 8 to 128GB (open dataset) →
- PossiblePossibly related (embedding) · 48%How're you deploying LLMs in production now-a-days? What's the best and most affordable way? [D] →
- PossiblePossibly related (embedding) · 50%Gemma 4 12B - MLX Kernel →
- PossiblePossibly related (embedding) · 46%CPU TTS benchmark with UTMOS MOS scoring: Kokoro, Supertonic, Inflect-Nano, and Kyutai's new Pocket TTS [P] →
- PossiblePossibly related (embedding) · 51%DynaMiCS: Fine-Tuning LLMs with Performance Constraints Using Dynamic Mixtures - Apple Machine Learning Research →
Covers
newsAsk HN: MacBook vs. Dedicated GPU for LLMnewsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsI mapped which local LLMs actually fit each RAM tier, 8 to 128GB (open dataset)newsHow're you deploying LLMs in production now-a-days? What's the best and most affordable way? [D]
Related to
Covers (incoming)
newsGemma 4 12B - MLX KernelnewsCPU TTS benchmark with UTMOS MOS scoring: Kokoro, Supertonic, Inflect-Nano, and Kyutai's new Pocket TTS [P]newsDynaMiCS: Fine-Tuning LLMs with Performance Constraints Using Dynamic Mixtures - Apple Machine Learning ResearchnewsApple M7 Ultra Chip Planned With Up to 1.5 TB of Unified MemorynewsApple M5 isn't making full use of its matmul cores yetnewsSOTA Apple Silicon Inference (August 15, 2026)
Related across the graph
newsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsApple M7 Ultra Chip Planned With Up to 1.5 TB of Unified MemorynewsDynaMiCS: Fine-Tuning LLMs with Performance Constraints Using Dynamic Mixtures - Apple Machine Learning ResearchnewsApple M5 isn't making full use of its matmul cores yetnewsI mapped which local LLMs actually fit each RAM tier, 8 to 128GB (open dataset)newsSOTA Apple Silicon Inference (August 15, 2026)newsCPU TTS benchmark with UTMOS MOS scoring: Kokoro, Supertonic, Inflect-Nano, and Kyutai's new Pocket TTS [P]newsHow're you deploying LLMs in production now-a-days? What's the best and most affordable way? [D]newsAsk HN: MacBook vs. Dedicated GPU for LLMtutorialEvaluate a model properlynewsGemma 4 12B - MLX Kernel
