Read original ↗
newsReddit r/LocalLLaMATrust 52 · CommunityPublished 28d agoLive · 27d ago

Fractale-350M-base: memory as trained behaviour instead of long context, a fully open research release

Some of you may remember my post about the research project behind this: a trained fast-weight memory, with the paper and the full research log at github.com/kkuette/thought-bank. This is the follow-up. The first public model of the series is out. Quick context: solo researcher, one RTX 3090 for everything below 97M params, openly working with Claude for implementation and write-ups. Direction and judgment are mine. What it is. Fra

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

Covers

Related across the graph