repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 9d ago
Menfre01/waveloom
为 DeepSeek 前缀缓存定制的终端 Code Agent(纯 Go),缓存命中率 95-99%,命中输入定价为未命中的 1/30。A terminal coding agent optimized for DeepSeek prefix caching — 95-99% cache hit, cache-hit input priced at 1/30th of cache-miss.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 51%DeepSeek-V4-Flash (MXFP4): compute buffer scales ~3x just from KV cache quant type (f16 vs q8_0) — anyone else seeing this? Llama.cpp →
- PossiblePossibly related (embedding) · 47%DeepSeek v4 Flash on 5090 in llama.cpp with 1 Million context →
- PossiblePossibly related (embedding) · 51%DeepSeek Releases V4 Pro With Higher Benchmarks, Open-Source Tooling, and Upcoming Price Increases - gHacks →
Covers
Covers (incoming)
Related across the graph
newsDeepSeek-V4-Flash (MXFP4): compute buffer scales ~3x just from KV cache quant type (f16 vs q8_0) — anyone else seeing this? Llama.cppnewsDeepSeek Releases V4 Pro With Higher Benchmarks, Open-Source Tooling, and Upcoming Price Increases - gHacksnewsDeepSeek v4 Flash on 5090 in llama.cpp with 1 Million context
