Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
ANGESTROM

The Intelligence Layer of Humanity. Everything AI. All in One Place.

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom Intelligence Private Limited. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /deepspeedai/DeepSpeed
Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 17h ago

deepspeedai/DeepSpeed

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • FuzzyOverlapping authors or contributors · 62%Pretraining Data Can Be Poisoned through Computational Propaganda →

    “Shared author/contributor keys: smith”

  • FuzzyOverlapping authors or contributors · 62%In-Place Tokenizer Expansion for Pre-trained LLMs →

    “Shared author/contributor keys: smith”

  • FuzzyOverlapping authors or contributors · 62%DSPrompt: Dynamic Soft Prompt Defense Against M-RAG Corruption →

    “Shared author/contributor keys: lai”

  • FuzzySimilar title/name (fuzzy) · 59%Calibration-Free Vehicle Speed Estimation: A Monocular Keypoint-Template Approach →

    “Fuzzy title match (0.73): “Calibration-Free Vehicle Speed Estimation: A Monocular Keypo” ≈ “deepspeedai/DeepSpeed””

  • FuzzyOverlapping authors or contributors · 62%Anchoring Instruction Outside Mask: Exact Reference Caching for Efficient In-Context Diffusion Transformers →

    “Shared author/contributor keys: lai”

  • PossiblePossibly related (embedding) · 49%DeepSeek open-sources inference optimizations with 60–85% faster generation [pdf] →
  • PossiblePossibly related (embedding) · 48%OpenAI and Broadcom unveil LLM-optimized inference chip →
  • PossiblePossibly related (embedding) · 47%HASTE: A Framework for Training-Free, Dynamic, and Steerable Compression of Pre-Trained Convolutional Neural Networks →

Implements

paperPretraining Data Can Be Poisoned through Computational PropagandapaperIn-Place Tokenizer Expansion for Pre-trained LLMspaperDSPrompt: Dynamic Soft Prompt Defense Against M-RAG CorruptionpaperCalibration-Free Vehicle Speed Estimation: A Monocular Keypoint-Template ApproachpaperAnchoring Instruction Outside Mask: Exact Reference Caching for Efficient In-Context Diffusion TransformerspaperHASTE: A Framework for Training-Free, Dynamic, and Steerable Compression of Pre-Trained Convolutional Neural NetworkspaperBeyond Adam: SOAP and Muon for Faster, Label-Efficient Training of Machine Learning Interatomic PotentialspaperDifferentiable Logic Gate Networks for Low-Latency EEG Classification on Edge DevicespaperEdgeBench: Unveiling Scaling Laws of Learning from Real-World EnvironmentspaperABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPUpaperHeadCast: Casting Attention Heads for Efficient Autoregressive Video GenerationpaperEfficient RLVR Scheduling via Graph-Structured Online Difficulty EstimationpaperAutonomous Agricultural Tractor: Integrated Weed Detection and LiDAR Navigation for Precision Paddy Farming

Covers

newsDeepSeek open-sources inference optimizations with 60–85% faster generation [pdf]newsOpenAI and Broadcom unveil LLM-optimized inference chipnewsWe’re launching the Google DeepMind Accelerator program in Asia Pacific to tackle environmental risks

contributed_to (incoming)

personjeffrapersonloadamspersontjruwasepersonstas00personmrwyattiipersontohtanapersonRezaYazdaniAminabadipersondelockpersonawan-10personconglonglipersonnelyahupersonlekurilepersonjomayeripersoninkcherrypersonsamyampersonYejing-Laipersoncmikeh2personcli99persondeepcharmpersonHeyangQinpersonrraminenpersonmolly-smithpersonsfc-gh-truwasepersonsamadejacobspersondigger-yu

Implements (incoming)

paperSystematic Evaluation of Learning Rate Scheduling Strategies Across Heterogeneous Architectures

Related across the graph

persontjruwasepersonawan-10personRezaYazdaniAminabadipersondelockpersondigger-yupaperAutonomous Agricultural Tractor: Integrated Weed Detection and LiDAR Navigation for Precision Paddy FarmingpersonrraminennewsOpenAI and Broadcom unveil LLM-optimized inference chippersonlekurilepersonsfc-gh-truwasepersonHeyangQinpaperDifferentiable Logic Gate Networks for Low-Latency EEG Classification on Edge DevicespersonYejing-LaipersonmrwyattiipaperBeyond Adam: SOAP and Muon for Faster, Label-Efficient Training of Machine Learning Interatomic Potentialspersoncli99paperPretraining Data Can Be Poisoned through Computational PropagandapaperIn-Place Tokenizer Expansion for Pre-trained LLMspaperAnchoring Instruction Outside Mask: Exact Reference Caching for Efficient In-Context Diffusion TransformerspaperEdgeBench: Unveiling Scaling Laws of Learning from Real-World EnvironmentspaperABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPUpaperCalibration-Free Vehicle Speed Estimation: A Monocular Keypoint-Template ApproachpaperHASTE: A Framework for Training-Free, Dynamic, and Steerable Compression of Pre-Trained Convolutional Neural NetworkspersondeepcharmpersonnelyahupersonsamadejacobspersonjomayeripaperDSPrompt: Dynamic Soft Prompt Defense Against M-RAG CorruptionpersonloadamsnewsWe’re launching the Google DeepMind Accelerator program in Asia Pacific to tackle environmental riskspersoncmikeh2personinkcherrypersonstas00persontohtanapersonjeffranewsDeepSeek open-sources inference optimizations with 60–85% faster generation [pdf]paperSystematic Evaluation of Learning Rate Scheduling Strategies Across Heterogeneous ArchitecturespersonsamyampaperHeadCast: Casting Attention Heads for Efficient Autoregressive Video GenerationpaperEfficient RLVR Scheduling via Graph-Structured Online Difficulty Estimationpersonmolly-smithpersonconglongli
Knowledge path··tjruwase→·awan-10→·RezaYazdaniAminabadi→Rdeepspeedai/DeepSpeed

Topics

billion-parameterscompressiondata-parallelismdeep-learninggpuinferencemachine-learningmixture-of-expertsmodel-parallelismpipeline-parallelism

Explore

Search similar →Knowledge graph →All repos →Full intelligence feed →
Maintenance94
RIS98GitHub verified
Graph trust82Primary
Graph score42997