repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago
engineering87/llm-atlas
Interactive, in-browser visualization of how a transformer language model works: tokens, attention, quantization, and sampling, rendered live.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- LinkedLinked via unknownTransformer →
- LinkedLinked via unknownIdentifying Interactions at Scale for LLMs →
- LinkedLinked via unknownmicrosoft/phi-2 →
- LinkedLinked via unknownsentence-transformers/all-MiniLM-L6-v2 →
- LinkedLinked via unknowndeepseek-ai/DeepSeek-V3 →
- LinkedLinked via unknownBook Review: Domain-Specific Small Language Models by Guglielmo Iozzia →
- LinkedLinked via unknownIEEE Rolls Out Large Language Models Virtual Training Course →
Related to
Covers
Covers (incoming)
newsBook Review: Domain-Specific Small Language Models by Guglielmo IozzianewsI shrank a transformer until every number fitted on the screen and made the weights editable [R]newsIEEE Rolls Out Large Language Models Virtual Training CoursenewsConlangCrafter Turns AI to Imagining LanguagesnewsKnowledge Distillation of Black-Box Large Language ModelsnewsKnowledge Distillation of Black-Box Large Language Models (2024)newsWhat performs the operations coordinated within each layer or head of a Transformer?newsMatching human intelligence might need more than LLMs - Transformer | SubstacknewsHow does a 102M-parameter transformer forecast multivariate time series?newsGerman SooFi team launches Soofi S 30B-A3B , an open-source Mixture-of-Experts (MoE) hybrid Mamba–Transformer foundation model for German and English.newsI built a compiler that turns computation graphs into the weights of a vanilla transformer — no training anywhere [P]newsPrismML Review: The Future of On-Device Large Language Models - quasa.io
Implements (incoming)
paperKnowledgeDebugger -- an Exploration Tool for Knowledge Localization and Editing in TransformerspaperThe State-Prediction Separation HypothesispaperTransformer Geometry Observatory TGO-II: Representational Similarity ObservatorypaperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPpaperInvariant Learning Dynamics of Transformers in Inductive Reasoning Tasks
Related across the graph
newsKnowledge Distillation of Black-Box Large Language Modelsmodelmicrosoft/phi-2paperThe State-Prediction Separation Hypothesisglossary_termTransformermodeldeepseek-ai/DeepSeek-V3paperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPnewsWhat performs the operations coordinated within each layer or head of a Transformer?newsI built a compiler that turns computation graphs into the weights of a vanilla transformer — no training anywhere [P]modelsentence-transformers/all-MiniLM-L6-v2paperTransformer Geometry Observatory TGO-II: Representational Similarity ObservatorypaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TasksnewsHow does a 102M-parameter transformer forecast multivariate time series?newsConlangCrafter Turns AI to Imagining LanguagesnewsMatching human intelligence might need more than LLMs - Transformer | SubstacknewsBook Review: Domain-Specific Small Language Models by Guglielmo IozzianewsI shrank a transformer until every number fitted on the screen and made the weights editable [R]newsKnowledge Distillation of Black-Box Large Language Models (2024)newsIEEE Rolls Out Large Language Models Virtual Training CoursepaperKnowledgeDebugger -- an Exploration Tool for Knowledge Localization and Editing in TransformersnewsGerman SooFi team launches Soofi S 30B-A3B , an open-source Mixture-of-Experts (MoE) hybrid Mamba–Transformer foundation model for German and English.newsPrismML Review: The Future of On-Device Large Language Models - quasa.ionewsIdentifying Interactions at Scale for LLMs
