Read original ↗
newsReddit r/LocalLLaMATrust 52 · CommunityPublished 1mo agoLive · 1mo ago

DeepSeek v4 Flash on 4090 + DDR5, my experience

Disclosure: No AI was used to write this My specs are: RTX 4090 128 GB DDR5 5600 MT/s Intel Core Ultra 7 270k Running nvidia-595 on ubuntu 26.04 with latest llama.cpp build (pulled and rebuilt this morning). Tried a lot of things, ended up running unsloth's UD-Q2_K_XL quant with command: taskset -c 0-7 /home/kevin/ai/llama.cpp/build/bin/llama-server -lv 4 -m /home/kevin/ai/models/DeepSeek-V4-Flash

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

Covers

Covers (incoming)

Related across the graph