repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · yesterday
benjaminzwhite/reasoning-models
Experiments with reasoning models, training techniques, papers
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 59%Northwind AI →
- PossiblePossibly related (embedding) · 50%The Riddle Riddle: Testing Flexible Reasoning in Large Language Models and Humans →
- PossiblePossibly related (embedding) · 49%Grounding LLM Reasoning under Incomplete Graph Evidence →
- PossiblePossibly related (embedding) · 48%Inference →
- PossiblePossibly related (embedding) · 48%Retrace-1.5B →
- PossiblePossibly related (embedding) · 50%Purified OPSD: On-Policy Self-Distillation Without Losing How to Think →
- PossiblePossibly related (embedding) · 45%Detecting Answer-Driven Reasoning in LLM-Based Educational Tutors via Truncated Chain-of-Thought Auditing →
- PossiblePossibly related (embedding) · 47%Knowledge Knows, Verbalization Tells: Disentangling Latent Directions for Mathematical Solvability in LLMs →
Related to
Implements
Implements (incoming)
paperPurified OPSD: On-Policy Self-Distillation Without Losing How to ThinkpaperDetecting Answer-Driven Reasoning in LLM-Based Educational Tutors via Truncated Chain-of-Thought AuditingpaperKnowledge Knows, Verbalization Tells: Disentangling Latent Directions for Mathematical Solvability in LLMspaperDT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety GuardrailpaperThink Through a Bottleneck: Hourglass Reasoning for Rigorous InductionpaperLLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with ElenchospaperLearning Mechanistic Reasoning for Chemical Reactions with Large Language Models
Related across the graph
paperLearning Mechanistic Reasoning for Chemical Reactions with Large Language ModelspaperDT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety GuardrailcompanyNorthwind AIpaperPurified OPSD: On-Policy Self-Distillation Without Losing How to ThinkpaperDetecting Answer-Driven Reasoning in LLM-Based Educational Tutors via Truncated Chain-of-Thought AuditingpaperGrounding LLM Reasoning under Incomplete Graph EvidencemodelRetrace-1.5BpaperThink Through a Bottleneck: Hourglass Reasoning for Rigorous InductionpaperKnowledge Knows, Verbalization Tells: Disentangling Latent Directions for Mathematical Solvability in LLMsglossary_termInferencepaperThe Riddle Riddle: Testing Flexible Reasoning in Large Language Models and HumanspaperLLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with Elenchos
