repoGitHubTrust 82 · PrimaryPublished 13d agoLive · 13d ago
Scottcjn/exo-cuda
Exo distributed inference with NVIDIA CUDA support via tinygrad
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 65%Alternative(s) to run CUDA on non-Nvidia hardware →
- PossiblePossibly related (embedding) · 56%Ubuntu, CUDA, llama.cpp , nvcc versioning →
- PossiblePossibly related (embedding) · 53%Run NVIDIA Nemotron and OpenAI GPT OSS models on Amazon Bedrock in AWS GovCloud (US) →
- PossiblePossibly related (embedding) · 51%Accelerating Block Low-Rank Foundation Model Inference on MemoryConstrained GPUs →
- PossiblePossibly related (embedding) · 54%Reduce ASR inference costs by 75% with NVIDIA MPS on Amazon EC2 →
Covers
Covers (incoming)
Related across the graph
newsNVIDIA dropped an NVIDIA-hosted CUDA MCP for AI-assisted CUDA operations, such as searching official, up-to-date documentation, writing optimized GPU code, and analyzing performance datanewsRun NVIDIA Nemotron and OpenAI GPT OSS models on Amazon Bedrock in AWS GovCloud (US)newsAccelerating Block Low-Rank Foundation Model Inference on MemoryConstrained GPUsnewsReduce ASR inference costs by 75% with NVIDIA MPS on Amazon EC2newsUbuntu, CUDA, llama.cpp , nvcc versioningnewsAlternative(s) to run CUDA on non-Nvidia hardware
