Read original ↗
newsReddit r/LocalLLaMATrust 52 · CommunityPublished 1mo agoLive · 1mo ago

I feel like I'm not using my hardware efficiently

Hi there, got a 7950x,128GB DDR5, RTX 4090 and RTX 3090TI. I'm currently running Qwen3.6 27B Q8 with 262k Context at Q8 with llama.cpp. It's not touching the DDR5 RAM at all but at the same time I couldn't get 122B A10B or the likes to run. Is my FOMO justified or isn't there anything better than this model to run currently? submitted by /u/Common

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

Covers (incoming)

Related across the graph