Skip to content
r/LocalLLaMA · Communities

Fractale-350M-base: memory as trained behaviour instead of long context, a fully open research release

Some of you may remember my post about the research project behind this: a trained fast-weight memory, with the paper and the full research log at github.com/kkuette/thought-bank. This is the follow-up. The first public model of the series is out. Quick context: solo researcher, one RTX 3090 for everything below 97M pa