paperarXivTrust 82 · PrimaryPublished 2mo agoLive · 2mo ago
Sparse attention at million-token context
A linear-cost attention variant that holds quality past a million tokens.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- LinkedLinked via unknownAttention →
- LinkedLinked via unknownlucidrains/poly-attention →
- LinkedLinked via unknownattention-zoo →
- LinkedLinked via unknownToken →
- LinkedLinked via unknownBreakthrough in long-context efficiency announced →
- LinkedLinked via unknownNew Server Hopes to Break Through AI’s “Memory Wall” →
- PossiblePossibly related (embedding) · 48%Looking for feedback on a small test SLM I built completely from scratch [P] →
- PossiblePossibly related (embedding) · 48%A Hippocampus for Linear Attention: An Exact Memory for What the Recurrent State Forgets →
Related to
Implements (incoming)
Related to (incoming)
Covers (incoming)
Related across the graph
newsNew Server Hopes to Break Through AI’s “Memory Wall”newsA Hippocampus for Linear Attention: An Exact Memory for What the Recurrent State ForgetsnewsLooking for feedback on a small test SLM I built completely from scratch [P]glossary_termTokenrepolucidrains/poly-attentionnewsBreakthrough in long-context efficiency announcedglossary_termAttentionrepoattention-zoo
