repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 24d ago
sgl-project/rbg
A workload for deploying LLM inference services on Kubernetes
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 51%Understanding dynamic resource allocation in Kubernetes →
- PossiblePossibly related (embedding) · 49%Implementing resilience patterns with Amazon Bedrock and LLM gateway →
- PossiblePossibly related (embedding) · 49%MCP Server Architecture Patterns for LLM-Integrated Applications →
- PossiblePossibly related (embedding) · 48%How're you deploying LLMs in production now-a-days? What's the best and most affordable way? [D] →
- PossiblePossibly related (embedding) · 45%Self-hosted GitHub Actions runners on Lambda MicroVMs →
- PossiblePossibly related (embedding) · 53%12 Ways to Reduce LLM Latency and Inference Costs in Production - KDnuggets →
- PossiblePossibly related (embedding) · 68%Running a self-hosted LLM in Kubernetes with vLLM →
Covers
Implements
Covers (incoming)
Related across the graph
paperMCP Server Architecture Patterns for LLM-Integrated ApplicationsnewsUnderstanding dynamic resource allocation in KubernetesnewsRunning a self-hosted LLM in Kubernetes with vLLMnewsSelf-hosted GitHub Actions runners on Lambda MicroVMsnewsImplementing resilience patterns with Amazon Bedrock and LLM gatewaynewsHow're you deploying LLMs in production now-a-days? What's the best and most affordable way? [D]news12 Ways to Reduce LLM Latency and Inference Costs in Production - KDnuggets
