X · @teortaxesTex
· X / Twitter
Does anyone have a clue of how much post-training compute is being used right now? V4-Pro is ≈1e25 class model (as are its peers). Over 2 months, cou…
Does anyone have a clue of how much post-training compute is being used right now? V4-Pro is ≈1e25 class model (as are its peers). Over 2 months, could they have spent another 1e25 on rollouts? More? What is the pretraining share at this point?