Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
Angestrom

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /amitshekhariitbhu/llm-internals
Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 12d ago

amitshekhariitbhu/llm-internals

Learn LLM internals step by step - from tokenization to attention to inference optimization.

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 51%Inference →
  • PossiblePossibly related (embedding) · 50%Understanding Large Language Models →
  • PossiblePossibly related (embedding) · 50%DemoPSD: Disagreement-Modulated Policy Self-Distillation →
  • PossiblePossibly related (embedding) · 50%DeepSeek open-sources inference optimizations with 60–85% faster generation [pdf] →
  • PossiblePossibly related (embedding) · 49%TOPS: First-Principles Visual Token Pruning via Constructing Token Optimal Preservation Sets for Efficient MLLM Inference →
  • PossiblePossibly related (embedding) · 48%I wrote a GGUF inferencer from scratch, AMA →
  • PossiblePossibly related (embedding) · 51%Estimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMs →
  • PossiblePossibly related (embedding) · 53%Towards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model Finetuning →

Related to

glossary_termInference

Implements

paperUnderstanding Large Language ModelspaperDemoPSD: Disagreement-Modulated Policy Self-DistillationpaperTOPS: First-Principles Visual Token Pruning via Constructing Token Optimal Preservation Sets for Efficient MLLM Inference

Covers

newsDeepSeek open-sources inference optimizations with 60–85% faster generation [pdf]

Covers (incoming)

newsI wrote a GGUF inferencer from scratch, AMA

Implements (incoming)

paperEstimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMspaperTowards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model FinetuningpaperThink Through a Bottleneck: Hourglass Reasoning for Rigorous InductionpaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperLLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with ElenchospaperExtractable Memorization From First PrinciplespaperLearning Mechanistic Reasoning for Chemical Reactions with Large Language ModelspaperKnowledgeless Language Models: Suppressing Parametric Recall for Evidence-Grounded Language Modeling

Related across the graph

paperLearning Mechanistic Reasoning for Chemical Reactions with Large Language ModelspaperDemoPSD: Disagreement-Modulated Policy Self-DistillationpaperEstimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMspaperTowards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model FinetuningpaperKnowledgeless Language Models: Suppressing Parametric Recall for Evidence-Grounded Language ModelingpaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperThink Through a Bottleneck: Hourglass Reasoning for Rigorous InductionpaperUnderstanding Large Language Modelsglossary_termInferencepaperExtractable Memorization From First PrinciplesnewsDeepSeek open-sources inference optimizations with 60–85% faster generation [pdf]newsI wrote a GGUF inferencer from scratch, AMApaperLLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with ElenchospaperTOPS: First-Principles Visual Token Pruning via Constructing Token Optimal Preservation Sets for Efficient MLLM Inference
Knowledge path·PLearning Mechanistic Reasoning for Chemical Reactions with Large Language Models→PDemoPSD: Disagreement-Modulated Policy Self-Distillation→PEstimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMs→Ramitshekhariitbhu/llm-internals

Topics

attention-is-all-you-needattention-mechanismlarge-language-modelslearn-llmllmllm-internals

Explore

Search similar →Knowledge graph →All repos →Full intelligence feed →
Graph trust82Primary
Graph score1480