Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
Angestrom

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /gpustack/gpustack
Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 2d ago

gpustack/gpustack

A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 65%Top Cost-Effective Enterprise GPU Cloud Platforms for AI Workloads with H100–GB200, Elastic Scaling and Pay-as-You-Go Compute - Scott Coop →
  • PossiblePossibly related (embedding) · 55%WattGPU: Predicting Inference Power and Latency on Unseen GPUs and LLMs →
  • PossiblePossibly related (embedding) · 52%NVIDIA Unlocks AI Compute at Scale, Inviting Partners to Power the AI Infrastructure Buildout →
  • PossiblePossibly related (embedding) · 52%SoftBank enters the rent-a-GPU race as America looks for support for AI training →
  • PossiblePossibly related (embedding) · 52%Claude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure →
  • PossiblePossibly related (embedding) · 52%Launch HN: machine0 (YC S26) – Persistent CPU and GPU VMs from the CLI →
  • PossiblePossibly related (embedding) · 66%I have a mid-sized GPU cluster and was thinking about giving free compute [D] →
  • PossiblePossibly related (embedding) · 55%Meet FreeToken: An Edge-Native MoE Serving Engine that Runs 753B GLM-5.2 on a Single Workstation GPU - MarkTechPost →

Covers

newsTop Cost-Effective Enterprise GPU Cloud Platforms for AI Workloads with H100–GB200, Elastic Scaling and Pay-as-You-Go Compute - Scott CoopnewsNVIDIA Unlocks AI Compute at Scale, Inviting Partners to Power the AI Infrastructure BuildoutnewsSoftBank enters the rent-a-GPU race as America looks for support for AI trainingnewsClaude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure

Implements

paperWattGPU: Predicting Inference Power and Latency on Unseen GPUs and LLMs

Covers (incoming)

newsLaunch HN: machine0 (YC S26) – Persistent CPU and GPU VMs from the CLInewsI have a mid-sized GPU cluster and was thinking about giving free compute [D]newsMeet FreeToken: An Edge-Native MoE Serving Engine that Runs 753B GLM-5.2 on a Single Workstation GPU - MarkTechPostnewsIntroducing @huggingface/kernels: 200+ WebGPU Kernels for Local AInewsShow HN: Computable – Buy, sell, and redeem GPU for the exact weeks you wantnewsGPUHedge: Hedging serverless GPU providers improves cold start p95 latency from 117s to 30s [P]

Implements (incoming)

paperTerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale

Related across the graph

paperTerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at ScalenewsIntroducing @huggingface/kernels: 200+ WebGPU Kernels for Local AInewsSoftBank enters the rent-a-GPU race as America looks for support for AI trainingnewsClaude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in AzurenewsGPUHedge: Hedging serverless GPU providers improves cold start p95 latency from 117s to 30s [P]paperWattGPU: Predicting Inference Power and Latency on Unseen GPUs and LLMsnewsMeet FreeToken: An Edge-Native MoE Serving Engine that Runs 753B GLM-5.2 on a Single Workstation GPU - MarkTechPostnewsLaunch HN: machine0 (YC S26) – Persistent CPU and GPU VMs from the CLInewsI have a mid-sized GPU cluster and was thinking about giving free compute [D]newsShow HN: Computable – Buy, sell, and redeem GPU for the exact weeks you wantnewsNVIDIA Unlocks AI Compute at Scale, Inviting Partners to Power the AI Infrastructure BuildoutnewsTop Cost-Effective Enterprise GPU Cloud Platforms for AI Workloads with H100–GB200, Elastic Scaling and Pay-as-You-Go Compute - Scott Coop
Knowledge path·PTerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale→NIntroducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI→NSoftBank enters the rent-a-GPU race as America looks for support for AI training→Rgpustack/gpustack

Topics

ascendcudadeepseekdistributed-inferencegenaihigh-performance-inferenceinferencellamallmllm-inference

Explore

Search similar →Knowledge graph →All repos →Full intelligence feed →
Graph trust82Primary
Graph score5573