r/LocalLLaMA
· Communities
bitsandbytes creator teasing new quantization method: GLM 5.3 on a single DGX Spark at 7t/s
Don't get too hyped + take with a grain of salt as there have been an endless amount of quantization schemes with big promises that never really became a thing. Tim Dettmers is a pretty well known researcher though, so maybe something will come of this. Time shall tell. Another tweet about the method, DS4 Pro on a sing