repoGitHubTrust 82 · PrimaryPublished 25d agoLive · 25d ago
bassrehab/triton-kernels
High-performance GPU kernels for LLM inference in OpenAI Triton. Fused RMSNorm, SwiGLU, INT8 GEMM with benchmarks and roofline analysis.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 54%[Paper] Automated Tensor Scheduling for Hybrid CPU-GPU LLM Inference on Consumer Devices →
- PossiblePossibly related (embedding) · 54%OpenAI and Broadcom unveil LLM-optimized inference chip →
- PossiblePossibly related (embedding) · 51%Accelerating Block Low-Rank Foundation Model Inference on MemoryConstrained GPUs →
- PossiblePossibly related (embedding) · 49%Bringing PyTorch Monarch to AMD GPUs →
