newsReddit r/LocalLLaMATrust 52 · CommunityPublished 12d agoLive · 12d ago
Qwen 3.8 27B Overthinking, It has to be done, it has to be overthinking to punch Opus 4.6
Yes, it sucks to waste time waiting on 16K+ reasoning tokens alone. But here's the thing, this is only a 27B model trying to perform on par with 1T+ parameter models. Something has to be sacrificed, and that sacrifice is the amount of reasoning or trajectory tokens. This isn't new to LLMs whatsoever. Andrej Karpathy himself has said that LLMs need tokens to think. He mentioned this somewhere in his "Let's build GPT" / GPT video series, although
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 57%pzqpzq/LSF_MDia →
- PossiblePossibly related (embedding) · 57%From Tokens to States: LLMs as a Special Case of World Models and the Continuous Path Beyond →
- PossiblePossibly related (embedding) · 54%Can We Break LLMs Out of Self-Loops? Fine-Grained Reasoning Control with Activation Steering →
- PossiblePossibly related (embedding) · 54%cmsa/manifest-driven-llms →
- PossiblePossibly related (embedding) · 54%The-Martyr/Awesome-Multimodal-Reasoning →
- PossiblePossibly related (embedding) · 47%jia-gao/leanctx →
Covers
Covers (incoming)
Related across the graph
repoThe-Martyr/Awesome-Multimodal-Reasoningrepocmsa/manifest-driven-llmsrepojia-gao/leanctxrepopzqpzq/LSF_MDiapaperFrom Tokens to States: LLMs as a Special Case of World Models and the Continuous Path BeyondpaperCan We Break LLMs Out of Self-Loops? Fine-Grained Reasoning Control with Activation Steering
