repoGitHubTrust 82 · PrimaryPublished 27d agoLive · 27d ago
microsoft/Tutel
Tutel MoE: Optimized Mixture-of-Experts Library, Support GptOss/DeepSeek/Kimi-K2/Qwen3 using FP8/NVFP4/MXFP4
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 54%H64LM: A 249M-parameter Mixture-of-Experts Transformer built from scratch in PyTorch [P] →
- PossiblePossibly related (embedding) · 49%DeepSeek open-sources inference optimizations with 60–85% faster generation [pdf] →
- PossiblePossibly related (embedding) · 48%Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains →
- PossiblePossibly related (embedding) · 48%Adaptive Mixture of Experts Gate (AMG) [R] →
