Skip to content
r/LocalLLaMA · Communities

Open labs are finally embracing the power of continued post training

A few years ago, every generation of model releases came entirely as new models, such as Qwen, Qwen 1.5, 2, 2.5, 3, and 3.5, llama, llama 2, llama 3, etc. But with the events of the recent few days, I think we can conclude that this is not the correct path forward. Continued post training on an existing model seems to