Skip to content
r/LocalLLaMA · Communities

Tested (the updated) Gemma 4 locally on coding with OpenCode

Gemma 4 was updated (mostly chat templates) and I took it for a test. On a local llama.cpp server running on M5 Pro with 48GB, 26B A4B (Q6) has about 60t/s and works well with OpenCode. It works quite well (given it's size) for backend work, but UI/UX is unacceptable. Watch the testing https://www.youtube.com/watch?v=m