Brrrrr 🚀 and it's free to use
Brrrrr 🚀and it's free to useqilua: glm 5.2 = 131 token/s 🙀
Brrrrr 🚀and it's free to useqilua: glm 5.2 = 131 token/s 🙀
RT HassanIntroducing The Blind Test.Two landing pages. One built by GLM 5.2 and one by Opus 4.8.Can you tell which is which?It's very difficult to get…
The next generation of inference needs purpose-built infrastructure.Together AI and 5C are deploying NVIDIA GB300 NVL72 systems with high-density compute, advanced cooling, and AI-optimized storage for…
A year ago this would have been an obvious closed-model task.Now GLM-5.2 can read the issue, reason through the scene, patch the code, and keep moving…
Voice agents get a lot more interesting when they can use the screen 🔥This demo runs the full loop on Together AI: STT, voice, and reasoning…
GLM-5.2 on Together AI is showing up fast on @OpenRouter ⚡️The model is strong, and our serving path makes that strength usable in the loop.Together has…
MiniMax-M3 expands what agents can carry into context: long histories, images, video, documents, and tool outputs.Together’s inference work makes that practical at scale by improving token…