repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago
Zefan-Cai/KVCache-Factory
Unified KV Cache Compression Methods for Auto-Regressive Models
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 46%GSRQ: Gain-Shape Residual Quantization for Sub-1-bit KV Cache →
- PossiblePossibly related (embedding) · 48%DepthWeave-KV: Token-Adaptive Cross-Layer Residual Factorization for Long-Context KV Cache Compression →
- PossiblePossibly related (embedding) · 45%SLORR: Simple and Efficient In-Training Low-Rank Regularization →
- PossiblePossibly related (embedding) · 55%A JoLT for the KV Cache: Near-Lossless KV Cache Compression via Joint Tucker and JL-Residual Allocation for LLMs →
- PossiblePossibly related (embedding) · 63%DKV: Open-source KV-cache compression framework for local LLM inference (CLI + technical report) →
Implements
Implements (incoming)
Covers (incoming)
Related across the graph
paperA JoLT for the KV Cache: Near-Lossless KV Cache Compression via Joint Tucker and JL-Residual Allocation for LLMspaperDepthWeave-KV: Token-Adaptive Cross-Layer Residual Factorization for Long-Context KV Cache CompressionpaperSLORR: Simple and Efficient In-Training Low-Rank RegularizationpaperGSRQ: Gain-Shape Residual Quantization for Sub-1-bit KV CachenewsDKV: Open-source KV-cache compression framework for local LLM inference (CLI + technical report)
