repoGitHubTrust 82 · PrimaryPublished 22d agoLive · 2h ago
defai-digital/ax-engine
One Mac process. Many models. Real speed. Multi-model LLM serving with prefix reuse, MTP acceleration, and OpenAI APIs — built for Apple Silicon, measured against mlx-lm and llama.cpp.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 51%Best Local LLMs You Can Run on a Single 24GB GPU in 2026: Qwen, Gemma, Mistral, DeepSeek Compared - MarkTechPost →
- PossiblePossibly related (embedding) · 51%Qualcomm launches GenieX to run LLMs on their Windows Laptops →
- PossiblePossibly related (embedding) · 48%OpenAI and Broadcom announce chip designed for LLM inference at scale →
- PossiblePossibly related (embedding) · 47%M5Stack Core2 gets open-source firmware to reproduce OpenAI’s Codex Micro features - CNX Software →
- PossiblePossibly related (embedding) · 47%Qwen vs Gemma →
Covers
newsBest Local LLMs You Can Run on a Single 24GB GPU in 2026: Qwen, Gemma, Mistral, DeepSeek Compared - MarkTechPostnewsQualcomm launches GenieX to run LLMs on their Windows LaptopsnewsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsM5Stack Core2 gets open-source firmware to reproduce OpenAI’s Codex Micro features - CNX SoftwarenewsQwen vs Gemma
Related across the graph
newsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsQwen vs GemmanewsBest Local LLMs You Can Run on a Single 24GB GPU in 2026: Qwen, Gemma, Mistral, DeepSeek Compared - MarkTechPostnewsQualcomm launches GenieX to run LLMs on their Windows LaptopsnewsM5Stack Core2 gets open-source firmware to reproduce OpenAI’s Codex Micro features - CNX Software
