newsReddit r/MachineLearningTrust 52 · CommunityPublished 1mo agoLive · 1mo ago
Chain of Thought is a scaling trap. the next wave is latent reasoning (Coconut / HRM / RecrusiveMAS)... but then we hit the black box wall. Where does BDH fit? [D]
Read a long piece on the future of LLM reasoning that makes a provocative claim: Chain of Thought is a useful hack but we've started to confuse a readable trace with the actual computation. All in all, "generating text is not the same as thinking." There are two practical problems here: Faithfulness: CoT style traces can decouple from what the model actually "did." u can get plausible steps with a wrong answer, or messy
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 63%Does Verbose Chain-of-Thought Really Help? In-Distribution Evidence that Content, Not Length, Matters →
- PossiblePossibly related (embedding) · 51%Bridging the Gap Between Latent and Explicit Reasoning with Looped Transformers →
- PossiblePossibly related (embedding) · 50%ThinkProbe: Beyond Accuracy -- Structural Profiling of Open-Ended LLM Reasoning Traces via Non-Generative Thought Graphs →
- PossiblePossibly related (embedding) · 50%CoLT: Teaching Multi-Modal Models to Think with Chain of Latent Thoughts →
- PossiblePossibly related (embedding) · 47%Message Passing Enables Efficient Reasoning →
- PossiblePossibly related (embedding) · 51%Training Continuous Chain of Thought Models: A Tale of Two Regimes →
- PossiblePossibly related (embedding) · 53%Token Budget Saturation and Mechanistic Early Detection of Reasoning Non-Convergence in Chain-of-Thought Models →
Covers
paperDoes Verbose Chain-of-Thought Really Help? In-Distribution Evidence that Content, Not Length, MatterspaperBridging the Gap Between Latent and Explicit Reasoning with Looped TransformerspaperThinkProbe: Beyond Accuracy -- Structural Profiling of Open-Ended LLM Reasoning Traces via Non-Generative Thought GraphspaperCoLT: Teaching Multi-Modal Models to Think with Chain of Latent ThoughtspaperMessage Passing Enables Efficient Reasoning
Covers (incoming)
Related across the graph
paperTraining Continuous Chain of Thought Models: A Tale of Two RegimespaperDoes Verbose Chain-of-Thought Really Help? In-Distribution Evidence that Content, Not Length, MatterspaperCoLT: Teaching Multi-Modal Models to Think with Chain of Latent ThoughtspaperMessage Passing Enables Efficient ReasoningpaperToken Budget Saturation and Mechanistic Early Detection of Reasoning Non-Convergence in Chain-of-Thought ModelspaperThinkProbe: Beyond Accuracy -- Structural Profiling of Open-Ended LLM Reasoning Traces via Non-Generative Thought GraphspaperBridging the Gap Between Latent and Explicit Reasoning with Looped Transformers
