Skip to content
r/LocalLLaMA · Communities

Qwen 3.6 27B – VLLM Performance Benchmark Results (BF16, FP8, NVFP4)

Sharing some testing of Qwen 3.6 27B using VLLM across the popular quants on my development system. I used llama benchy to generate the results, then fed it into an LLM to format it the tables for readibility. While NVFP4 is blazing fast, have had looping issues in copilot that I don't get with BF16, and the responses