X · @teortaxesTex
· X / Twitter
Goes to show that Chinese models are still undertrained and don't have enough RL. Computationally modest efforts can still push them harder. Should re…
Goes to show that Chinese models are still undertrained and don't have enough RL. Computationally modest efforts can still push them harder. Should reduce your prior on the utility of large-scale distillation.Lentils: Macaron V1 Venti, the first model to be post-trained on GLM-5.2, is released todayIt seems to be a dec