Skip to content
X · @teortaxesTex · X / Twitter

shitposting as a post-training paradigm

shitposting as a post-training paradigmBasalt: Introducing Monolith-1.0: Frontier Reasoning, Open Weights- 1.57T-param Mixture-of-Experts, 49.5B active per token- Native 1M-token context (2²⁰), extended via a two-stage YaRN curriculum- Trained on 60T tokens across 12,288 Ascend 910C NPUs- Frontier SOTA open-source mode