repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 7d ago
alibaba/MNN
MNN: A blazing-fast, lightweight inference engine battle-tested by Alibaba, powering high-performance on-device LLMs and Edge AI.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 66%OpenAI and Broadcom unveil LLM-optimized inference chip →
- PossiblePossibly related (embedding) · 62%OpenAI and Broadcom announce chip designed for LLM inference at scale →
- PossiblePossibly related (embedding) · 57%Hardware startup unveils inference accelerator →
- PossiblePossibly related (embedding) · 53%Cerebras OpenAI deal capacity has effectively killed the waitlist for everyone else [D] →
- PossiblePossibly related (embedding) · 52%OpenAI reveals its first AI processor: Jalapeño →
- PossiblePossibly related (embedding) · 47%FastFlowLM Joins AMD to Advance AI Inference →
- PossiblePossibly related (embedding) · 26%SelectInfer: Selective Neuron Loading and Computation for On-Device LLMs →
“Possibly related via embedding similarity 0.57 (not asserted). Timestamp check: artifact slightly before paper (-14d).”
- PossiblePossibly related (embedding) · 54%High-Performance MoE Inference: Qwen3.6–35B-A3B on an AI PC with OpenVINO - Medium →
Covers
newsOpenAI and Broadcom unveil LLM-optimized inference chipnewsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsHardware startup unveils inference acceleratornewsCerebras OpenAI deal capacity has effectively killed the waitlist for everyone else [D]newsOpenAI reveals its first AI processor: Jalapeño
Covers (incoming)
newsFastFlowLM Joins AMD to Advance AI InferencenewsHigh-Performance MoE Inference: Qwen3.6–35B-A3B on an AI PC with OpenVINO - MediumnewsOpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks shownewsJalapeño’s first results show industry-leading speed and efficiency in AI inferencenewsConjure cash with old Macs by linking them to AI inference BorgnewsOpenAI's upcoming Jalapeño chip looks like it'll be an inference beast
Related to (incoming)
Related across the graph
newsOpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks showpaperLlama-Mobile: Efficient 2.7-Bit Quantization of VLMspaperSelectInfer: Selective Neuron Loading and Computation for On-Device LLMsnewsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsCerebras OpenAI deal capacity has effectively killed the waitlist for everyone else [D]newsOpenAI and Broadcom unveil LLM-optimized inference chipnewsConjure cash with old Macs by linking them to AI inference BorgnewsFastFlowLM Joins AMD to Advance AI InferencenewsHigh-Performance MoE Inference: Qwen3.6–35B-A3B on an AI PC with OpenVINO - MediumnewsOpenAI reveals its first AI processor: JalapeñonewsHardware startup unveils inference acceleratornewsOpenAI's upcoming Jalapeño chip looks like it'll be an inference beastpaperPre-Compiled Pipeline Shards for Distributed LLM Inference on Intel AI PC FleetsnewsJalapeño’s first results show industry-leading speed and efficiency in AI inference
