Skip to content
X · @elonmusk · X / Twitter

RT DogeDesigner: BREAKING: Grok 4.5 (high) ranks #1 on the HighWalk benchmark, which tests how well AI agents update technical specifications from cod…

RT DogeDesignerBREAKING: Grok 4.5 (high) ranks #1 on the HighWalk benchmark, which tests how well AI agents update technical specifications from code changes.Grok delivered the best combination of quality and operational efficiency, finishing ahead of Claude and GPT.