Skip to content
r/LocalLLaMA · Communities

Qwen3.6-27B: NVFP4/FP8 agent loops vs flawless BF16. Config or quant issue?

Hi everyone, I'm trying to determine if I'm dealing with a misconfiguration in my stack or if this is an inherent limitation of current quantization methods for agentic workflows. I recently set up a dedicated rig with an RTX PRO 6000 Blackwell and have been benchmarking Qwen3.6-27B, but I'm hitting severe reliability