Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
Angestrom

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /InternLM/xtuner
Read original ↗
repoGitHubTrust 82 · PrimaryPublished 2mo agoLive · 5d ago

InternLM/xtuner

A Next-Generation Training Engine Built for Ultra-Large MoE Models

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 50%Qwen3.6-27B UD Q3 with kv at q8 is quite amazing for simple proof of concepts →
  • PossiblePossibly related (embedding) · 47%Meituan Open-Sources 1.6-Trillion-Parameter LongCat-2.0, First Model Fully Trained on 50,000 Chinese-Made AI Chips - finance.biggo.com →
  • PossiblePossibly related (embedding) · 45%Tencent-HY3 is the real deal on 128GB! →
  • PossiblePossibly related (embedding) · 50%Benchmarks: AntLing-3.0-flash a hybrid-reasoning MoE model built for production-scale agents. →
  • PossiblePossibly related (embedding) · 49%Don't want to be this guy, but I need Qwen 3.8 35B A3B →
  • PossiblePossibly related (embedding) · 51%Proposed architecture for inferencing sparse MOE models increasing Active parameters using layered + linear decay. Succinct reasoning without any model training or fine tune. [p] →

Covers

newsQwen3.6-27B UD Q3 with kv at q8 is quite amazing for simple proof of conceptsnewsMeituan Open-Sources 1.6-Trillion-Parameter LongCat-2.0, First Model Fully Trained on 50,000 Chinese-Made AI Chips - finance.biggo.comnewsTencent-HY3 is the real deal on 128GB!

Covers (incoming)

newsBenchmarks: AntLing-3.0-flash a hybrid-reasoning MoE model built for production-scale agents.newsDon't want to be this guy, but I need Qwen 3.8 35B A3BnewsProposed architecture for inferencing sparse MOE models increasing Active parameters using layered + linear decay. Succinct reasoning without any model training or fine tune. [p]

Related across the graph

newsProposed architecture for inferencing sparse MOE models increasing Active parameters using layered + linear decay. Succinct reasoning without any model training or fine tune. [p]newsDon't want to be this guy, but I need Qwen 3.8 35B A3BnewsBenchmarks: AntLing-3.0-flash a hybrid-reasoning MoE model built for production-scale agents.newsMeituan Open-Sources 1.6-Trillion-Parameter LongCat-2.0, First Model Fully Trained on 50,000 Chinese-Made AI Chips - finance.biggo.comnewsTencent-HY3 is the real deal on 128GB!newsQwen3.6-27B UD Q3 with kv at q8 is quite amazing for simple proof of concepts
Knowledge path·NProposed architecture for inferencing sparse MOE models increasing Active parameters using layered + linear decay. Succinct reasoning without any model training or fine tune. [p]→NDon't want to be this guy, but I need Qwen 3.8 35B A3B→NBenchmarks: AntLing-3.0-flash a hybrid-reasoning MoE model built for production-scale agents.→RInternLM/xtuner

Topics

agentdeepseek-v3gpt-ossintern-s1internvlkimi-k2llmmultimodalqwen3-moeqwen3-vl

Explore

Search similar →Knowledge graph →All repos →Full intelligence feed →
Graph trust82Primary
Graph score5194