Skip to content
r/LocalLLaMA · Communities

bitsandbytes creator teasing new quantization method: GLM 5.3 on a single DGX Spark at 7t/s

Don't get too hyped + take with a grain of salt as there have been an endless amount of quantization schemes with big promises that never really became a thing. Tim Dettmers is a pretty well known researcher though, so maybe something will come of this. Time shall tell. Another tweet about the method, DS4 Pro on a sing