Skip to content
r/LocalLLaMA · Communities

GLM 5.2 running on MacBook Pro M5 48 GB Ram at between 2 – 2.8t/s

I started reading about Flash MOE I have my own built Claude Desktop style app using Pi as the harness. Qwen3.6 27B is good but sometimes it falls short on some of the large codebases I work on so I wanted to see if I could get GLM 5.2 working on my machine with Flash MOE. Got to work with Claude and current benchmarks