repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · yesterday
jundot/omlx
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 57%OpenAI and Broadcom announce chip designed for LLM inference at scale →
- PossiblePossibly related (embedding) · 50%Hardware startup unveils inference accelerator →
- PossiblePossibly related (embedding) · 49%OpenAI and Broadcom unveil LLM-optimized inference chip →
- PossiblePossibly related (embedding) · 49%I mapped which local LLMs actually fit each RAM tier, 8 to 128GB (open dataset) →
- PossiblePossibly related (embedding) · 50%Llama-Server is Throwing Away Your Perfectly Good KV Caches, and How to Fix It →
- PossiblePossibly related (embedding) · 46%Elastic Gang: Per-Token Membership Change for a Hard-Barriered LLM Inference Gang Co-Scheduled with OS Processes →
- PossiblePossibly related (embedding) · 52%CachyLLama: llama.cpp fork with persistent SSD-backed KV caching for local agent workflows →
Covers
Covers (incoming)
Implements (incoming)
Related across the graph
newsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsOpenAI and Broadcom unveil LLM-optimized inference chipnewsI mapped which local LLMs actually fit each RAM tier, 8 to 128GB (open dataset)paperElastic Gang: Per-Token Membership Change for a Hard-Barriered LLM Inference Gang Co-Scheduled with OS ProcessesnewsCachyLLama: llama.cpp fork with persistent SSD-backed KV caching for local agent workflowsnewsLlama-Server is Throwing Away Your Perfectly Good KV Caches, and How to Fix ItnewsHardware startup unveils inference accelerator
