GLM 5.2 on consumer hardware
I tried out the unsloth quants of GLM 5.2 on still "consumer-ish" hardware: 32C Zen5 Threadripper Pro 9975 WX, Asus WRX90E-SAGE-SE PCIe Gen5, 512GB DDR5 ECC…
I tried out the unsloth quants of GLM 5.2 on still "consumer-ish" hardware: 32C Zen5 Threadripper Pro 9975 WX, Asus WRX90E-SAGE-SE PCIe Gen5, 512GB DDR5 ECC…
The idea that distilling from Opus 4.8 lets you reach Mythos is very encouraging. It would mean that some GLM 5.3 would be enough to do…
Takeout delivery planning evolution over GLM-5 series.This is a really creative (and very Chinese) benchmark(Meituan sweating bullets rn)karminski-牙医: 聊聊智谱市值破万亿为什么不是高估事先声明, 个人观点仅供参考. 直接说结论, 智谱在 GLM 的 Agent 能力训练上是有东西的.…
submitted by /u/Intrepid_Rub_3566 [link] [comments]
Some of you may remember old models from GLM: GLM Air or GLM Flash. I know they’re outdated, but I have a soft spot for them,…
RT Dmytro Dzhulgakovyou may have heard that glm-5.2 at 280 token/s is cool, how about 318and we still have room to go
Cost per iteration is the unlock.GLM-5.2 on Together AI can generate polished web apps for a few cents. At that price, developers can explore more directions,…
Hello guys, hoping you're doing fine! I was wondering, for users with 4x-8x 6000 PROs (so between 384 and 768GB VRAM), how are bigger models working…
RT Zixuan LiRight alongside Cursor, Devin Desktop (Windsurf) and CLI now support GLM-5.2 as well.FrontierCode Extended is a benchmark we care deeply about for real-world engineering…
RT Zixuan LiGLM-5.2 is now available in Cursor.The model has performed strongly on OpenRouter's Cursor usage rankings over the past week. Would love to hear comparisons…
GLM 5.2 being on the Opus frontier for cost of CursorBench is what drives frontier lab margins downLee Robinson: You can now try GLM 5.2 in…
RT Vipul Ved PrakashA tangible comparison of GLM performance stacked against Opus 4.8 on web tasks by @nutlope. GLM 5.2 is chattier, but still faster when…
TL;DR: the recipe's image-build mods aren't actually public – I reconstructed them from the public kernels (with Claude) – and you have to build vLLM at…
Add more wins for GLM.The model has some brittle characteristics, and is getting crushed by closed models here, but we should expect open models to be…
GLM 5.2 is the best Chinese model on ARC-AGI-2, at 22.8% (is that high or max?), on par with Opus 4.5 (16K). …Whereas Grok 4.20 is…
Zhipu AI's GLM-5.2 nearly matches Claude Opus 4.7 in a Snowflake benchmark with 103 coding tasks at one-fifth the cost per output token. But the Chinese…
RT HassanAnnouncing GLM Arena!A series of tests (infographics, svgs, sites, ect..) ran on GLM 5.2 and Opus 4.8, with prompts included.On average, GLM 5.2 produced 2x…
G'day. This is part 3 on my Local LLM adventures. I have a crazy system hacked server-to-desktop system: Component Spec GPUs 2x Hopper H100, 96 GB…
btw Zai IPO'ed in Jan at HK$120 a share. when I first met @louszbd nobody really knew anyone using GLM's.now they have beat deepseek with the…
RT HassanRan 10 more tests comparing GLM 5.2 & Opus.On average, GLM 5.2 produced 2x the tokens but was still faster + 3x cheaper with similar…
Hey all, I have a piece of hardware laying around which is pretty fast from a traditional (non-GPU) server viewpoint. The hardware is the following: Dell…
We build a practical GLM-5.2 workflow using its hosted, OpenAI-compatible API instead of running the model locally. We set up multiple providers, load the API key…
RT Zixuan LiGLM-5.2 is available in Perplexity's Agent API. Just tested it, and it's powerful when paired with the Search SDK inside a sandbox.- Spin up…
RT afra wangthing i know about http://z.ai, the company behind GLM:1. http://z.ai is known as "Zhipu" before being rebranded as a sleeker "http://z.ai". The Chinese name…
GLM is Zhipu AI's open-weight model line — ChatGLM, GLM-4, GLM-4.5, GLM-V (vision). One of the earliest Chinese open-weight families; the 4.5 generation competes with the strongest Western models on agentic benchmarks.
Owner: Zhipu. We have 235 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is chatglm.cn.
Related text models: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen.