repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago
rasbt/reasoning-from-scratch
Implement a reasoning LLM in PyTorch from scratch, step by step
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 53%I wrote a GGUF inferencer from scratch, AMA →
- PossiblePossibly related (embedding) · 52%H64LM: A 249M-parameter Mixture-of-Experts Transformer built from scratch in PyTorch [P] →
- PossiblePossibly related (embedding) · 48%Profiling in PyTorch (Part 2): From nn.Linear to a Fused MLP →
- PossiblePossibly related (embedding) · 48%ThinkProbe: Beyond Accuracy -- Structural Profiling of Open-Ended LLM Reasoning Traces via Non-Generative Thought Graphs →
- PossiblePossibly related (embedding) · 48%COCOLogic-V2: Identifying Logical Inconsistencies via Truly Hard-Negatives →
- FuzzySimilar title/name (fuzzy) · 59%Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models →
“Fuzzy title match (0.73): “Towards Enhancing 3D Spatial Reasoning in Medical Multimodal” ≈ “rasbt/reasoning-from-scratch””
- FuzzySimilar title/name (fuzzy) · 59%Learning Mechanistic Reasoning for Chemical Reactions with Large Language Models →
“Fuzzy title match (0.73): “Learning Mechanistic Reasoning for Chemical Reactions with L” ≈ “rasbt/reasoning-from-scratch””
- FuzzySimilar title/name (fuzzy) · 59%SCOPE-RL: Optimizing Reasoning Paths Before and After Success →
“Fuzzy title match (0.73): “SCOPE-RL: Optimizing Reasoning Paths Before and After Succes” ≈ “rasbt/reasoning-from-scratch””
Covers
Implements
paperThinkProbe: Beyond Accuracy -- Structural Profiling of Open-Ended LLM Reasoning Traces via Non-Generative Thought GraphspaperCOCOLogic-V2: Identifying Logical Inconsistencies via Truly Hard-NegativespaperTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelspaperLearning Mechanistic Reasoning for Chemical Reactions with Large Language ModelspaperSCOPE-RL: Optimizing Reasoning Paths Before and After SuccesspaperLearning to Evolve Scenes: Reasoning about Human Activities with Scene Graphspaper4DR360: State Reasoning for Joint 3D Detection and Occupancy Prediction in 4D Radar-Camera Full-Scene PerceptionpaperDiffusion-GR2: Diffusion Generative Reasoning Re-rankerpaperReContext: Recursive Evidence Replay as LLM Harness for Long-Context ReasoningpaperWILDTRACE: Benchmarking Natural Evidence Trails in Long-Context ReasoningpaperMulti-scale Object-Aware Gaze Estimation via Geometric ReasoningpaperToken-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement LearningpaperG-RRM: Guiding Symbolic Solvers with Recurrent Reasoning ModelspaperBefore Thinking, Learn to Decide: Proactive Routing for Efficient Visual ReasoningpaperVisual Access Boundaries in Vision-Language Model ReasoningpaperEstimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMspaperThe Complexity Ceiling Benchmark: A Multi-Domain Evaluation of Sequential Reasoning Under Depth ScalingpaperVerifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage ControlpaperIdeas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea GenerationpaperDT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety GuardrailpaperCLExEval: A Human-in-the-Loop Framework for Qualitative Evaluation of LLM Clinical ReasoningpaperAre We Measuring Strategy or Phrasing? The Gap Between Surface- and Approach-Level Diversity in LLM Math ReasoningpaperEchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in EchocardiographypaperTask-Specific Multimodal Question Answering Agents via Confidence Calibration and Incremental Reasoning for QANTA 2026paperDo Vision-Language-Action Models Mean What They Say? On the Role of Faithfulness in Embodied ReasoningpaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperReaORE: Reasoning-Guided Progressive Open Relation Extraction Empowered by Large Reasoning ModelspaperDeep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning ModelspaperShow Me How You Reason and I'll Tell You Who You Are: Reasoning Graphs for Robust LLM Authorship AttributionpaperCoTu at EXACT 2026: Neuro-Symbolic Reasoning for Transparent Educational QApaperGold-Guided Programmatic Distillation for Financial Reasoning over Hybrid Tables and TextpaperStop Thinking, Start Looking: Efficient Post-Training for Multimodal Document Question Answering via Reasoning-Free AlignmentpaperLeveraging Instruction Tuning and Merging for Reasoning Model AdaptationpaperHierarchical Denoising For Multi-Step Visual ReasoningpaperHoloGeo: Mitigating Landmark Bias in Geo-localization via Evidence-Driven ReasoningpaperDetecting Answer-Driven Reasoning in LLM-Based Educational Tutors via Truncated Chain-of-Thought AuditingpaperLearning to Select, Not Relearn: Hard-Routed Mixtures of Reasoning LoRAspaperDo AI Agents Know When a Task Is Simple? Toward Complexity-Aware Reasoning and ExecutionpaperA rubric-based controlled comparison of frontier language models on expert-authored clinical reasoning tasks
Covers (incoming)
Implements (incoming)
Related to (incoming)
Related across the graph
paper4DR360: State Reasoning for Joint 3D Detection and Occupancy Prediction in 4D Radar-Camera Full-Scene PerceptionpaperDiffusion-GR2: Diffusion Generative Reasoning Re-rankerpaperSCOPE-RL: Optimizing Reasoning Paths Before and After SuccesspaperReContext: Recursive Evidence Replay as LLM Harness for Long-Context ReasoningnewsProfiling in PyTorch (Part 2): From nn.Linear to a Fused MLPpaperLearning Mechanistic Reasoning for Chemical Reactions with Large Language ModelspaperVerifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage ControlpaperWILDTRACE: Benchmarking Natural Evidence Trails in Long-Context ReasoningpaperMulti-scale Object-Aware Gaze Estimation via Geometric ReasoningpaperToken-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement LearningpaperCOCOLogic-V2: Identifying Logical Inconsistencies via Truly Hard-NegativespaperG-RRM: Guiding Symbolic Solvers with Recurrent Reasoning ModelspaperLearning to Evolve Scenes: Reasoning about Human Activities with Scene GraphspaperBefore Thinking, Learn to Decide: Proactive Routing for Efficient Visual ReasoningpaperHierarchical Denoising For Multi-Step Visual ReasoningpaperVisual Access Boundaries in Vision-Language Model ReasoningnewsProfiling in PyTorch (Part 3): Attention is all you profilepaperEstimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMspaperThe Complexity Ceiling Benchmark: A Multi-Domain Evaluation of Sequential Reasoning Under Depth ScalingpaperHoloGeo: Mitigating Landmark Bias in Geo-localization via Evidence-Driven ReasoningpaperIdeas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea GenerationpaperDT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety GuardrailpaperAre We Measuring Strategy or Phrasing? The Gap Between Surface- and Approach-Level Diversity in LLM Math ReasoningpaperStop Thinking, Start Looking: Efficient Post-Training for Multimodal Document Question Answering via Reasoning-Free AlignmentpaperCLExEval: A Human-in-the-Loop Framework for Qualitative Evaluation of LLM Clinical ReasoningpaperDo AI Agents Know When a Task Is Simple? Toward Complexity-Aware Reasoning and ExecutionpaperDo Vision-Language-Action Models Mean What They Say? On the Role of Faithfulness in Embodied ReasoningpaperEchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in EchocardiographypaperDetecting Answer-Driven Reasoning in LLM-Based Educational Tutors via Truncated Chain-of-Thought AuditingpaperLearning to Select, Not Relearn: Hard-Routed Mixtures of Reasoning LoRAspaperDeep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning ModelspaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TasksmodelJackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-DistilledpaperReaORE: Reasoning-Guided Progressive Open Relation Extraction Empowered by Large Reasoning ModelspaperTask-Specific Multimodal Question Answering Agents via Confidence Calibration and Incremental Reasoning for QANTA 2026paperA rubric-based controlled comparison of frontier language models on expert-authored clinical reasoning taskspaperThink Through a Bottleneck: Hourglass Reasoning for Rigorous InductionnewsH64LM: A 249M-parameter Mixture-of-Experts Transformer built from scratch in PyTorch [P]paperThinkProbe: Beyond Accuracy -- Structural Profiling of Open-Ended LLM Reasoning Traces via Non-Generative Thought GraphspaperLeveraging Instruction Tuning and Merging for Reasoning Model AdaptationpaperTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelspaperShow Me How You Reason and I'll Tell You Who You Are: Reasoning Graphs for Robust LLM Authorship AttributionpaperGold-Guided Programmatic Distillation for Financial Reasoning over Hybrid Tables and TextpaperCoTu at EXACT 2026: Neuro-Symbolic Reasoning for Transparent Educational QAnewsI wrote a GGUF inferencer from scratch, AMA
