newsReddit r/LocalLLaMATrust 52 · CommunityPublished 23d agoLive · 23d ago
A caveman qwen3.6 27B
Just saw this on huggingface: https://huggingface.co/ProCreations/grug-27b The benchmarks claim that it's quite a bit better than qwen3.6 27B original and that they reduced the amount of necessary tokens by more than 90%. It would make 27B running on my old laptop at 3tps feel more like 30tps for the thinking part, if true. Couldn't test it yet. submitted by &
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 48%huggingface/chat-ui →
- PossiblePossibly related (embedding) · 47%Qwen/QwQ-32B →
