Skip to content
r/LocalLLaMA · Communities

Ternary Qwen3.6 27B Tested on 3090!

I can how run 60 tk/s with two slot now, quality seems good, tool call is very stable. I haven't done any coding yet. 2 slot each have 100k KV cache allocated and it took around 21GB of VRAM submitted by /u/Top_Outlandishness78 [link] [comments]