newsGoogle News — Machine LearningTrust 62 · AggregatorPublished 1mo agoLive · 1mo ago
Transformers in Deep Learning: How Self-Attention Changed Modern AI - Snowflake
Transformers in Deep Learning: How Self-Attention Changed Modern AI Snowflake
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 56%Understanding Large Language Models →
- PossiblePossibly related (embedding) · 52%huggingface/transformers →
- PossiblePossibly related (embedding) · 50%Transformer →
- PossiblePossibly related (embedding) · 50%Transformer Geometry Observatory TGO-II: Representational Similarity Observatory →
- PossiblePossibly related (embedding) · 50%darkdevil3610/100-AI-Machine-learning-Deep-learning-Computer-vision-NLP →
- PossiblePossibly related (embedding) · 51%alternbits/awesome-ai-newsletters →
- PossiblePossibly related (embedding) · 56%From Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASP →
- PossiblePossibly related (embedding) · 50%aymericdamien/TopDeepLearning →
Covers
Covers (incoming)
paperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPrepoaymericdamien/TopDeepLearningpaperInhibited Self-Attention: Sharpening Focus in Vision TransformerspaperToward Localizing and Repairing Bias in Transformer Attention HeadspaperELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training TransformerspaperInvariant Learning Dynamics of Transformers in Inductive Reasoning Tasksreporoatienza/Deep-Learning-ExperimentspaperMobius Learning: Cyclic Depth Folding in Transformers
Related across the graph
paperInhibited Self-Attention: Sharpening Focus in Vision TransformerspaperToward Localizing and Repairing Bias in Transformer Attention Headsglossary_termTransformerpaperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPpaperELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training Transformersrepodarkdevil3610/100-AI-Machine-learning-Deep-learning-Computer-vision-NLPpaperTransformer Geometry Observatory TGO-II: Representational Similarity ObservatorypaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperMobius Learning: Cyclic Depth Folding in TransformerspaperUnderstanding Large Language Modelsrepoalternbits/awesome-ai-newslettersrepoaymericdamien/TopDeepLearningreporoatienza/Deep-Learning-Experimentsrepohuggingface/transformers
