Skip to content
r/LocalLLaMA · Communities

CTX: How far can you reasonably go with Qwen 3.6 27B?

How far can i stretch the context window with Qwen 3.6 27B (using Q8_0) before it gets too unreliable? I am at 100k right now and i am not quite statisfied. Other than not quantizing KV cache, is there anything else that can be done to make the model more stable over longer CTX? submitted by /u/milpster [link] [comment