repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 12d ago
amitshekhariitbhu/llm-internals
Learn LLM internals step by step - from tokenization to attention to inference optimization.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 51%Inference →
- PossiblePossibly related (embedding) · 50%Understanding Large Language Models →
- PossiblePossibly related (embedding) · 50%DemoPSD: Disagreement-Modulated Policy Self-Distillation →
- PossiblePossibly related (embedding) · 50%DeepSeek open-sources inference optimizations with 60–85% faster generation [pdf] →
- PossiblePossibly related (embedding) · 49%TOPS: First-Principles Visual Token Pruning via Constructing Token Optimal Preservation Sets for Efficient MLLM Inference →
- PossiblePossibly related (embedding) · 48%I wrote a GGUF inferencer from scratch, AMA →
- PossiblePossibly related (embedding) · 51%Estimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMs →
- PossiblePossibly related (embedding) · 53%Towards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model Finetuning →
Related to
Implements
Covers
Covers (incoming)
Implements (incoming)
paperEstimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMspaperTowards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model FinetuningpaperThink Through a Bottleneck: Hourglass Reasoning for Rigorous InductionpaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperLLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with ElenchospaperExtractable Memorization From First PrinciplespaperLearning Mechanistic Reasoning for Chemical Reactions with Large Language ModelspaperKnowledgeless Language Models: Suppressing Parametric Recall for Evidence-Grounded Language Modeling
Related across the graph
paperLearning Mechanistic Reasoning for Chemical Reactions with Large Language ModelspaperDemoPSD: Disagreement-Modulated Policy Self-DistillationpaperEstimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMspaperTowards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model FinetuningpaperKnowledgeless Language Models: Suppressing Parametric Recall for Evidence-Grounded Language ModelingpaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperThink Through a Bottleneck: Hourglass Reasoning for Rigorous InductionpaperUnderstanding Large Language Modelsglossary_termInferencepaperExtractable Memorization From First PrinciplesnewsDeepSeek open-sources inference optimizations with 60–85% faster generation [pdf]newsI wrote a GGUF inferencer from scratch, AMApaperLLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with ElenchospaperTOPS: First-Principles Visual Token Pruning via Constructing Token Optimal Preservation Sets for Efficient MLLM Inference
