Skip to content
X · @elonmusk · X / Twitter

RT DogeDesigner: Grok 4.6 ranks #1 on RuntimeWire’s Newsroom Reliability v0.2 benchmark with a score of 0.79, beating GPT-5.6 Sol, Claude Opus 4.8, G…

RT DogeDesignerGrok 4.6 ranks #1 on RuntimeWire’s Newsroom Reliability v0.2 benchmark with a score of 0.79, beating GPT-5.6 Sol, Claude Opus 4.8, Gemini and DeepSeek.