you can now fine tune Prism-ML’s ternary Bonsai models
examples included, use a high damn learning rate: https://github.com/electroglyph/ternary_QAT submitted by /u/terminoid_ [link] [comments]
examples included, use a high damn learning rate: https://github.com/electroglyph/ternary_QAT submitted by /u/terminoid_ [link] [comments]
submitted by /u/Mochila-Mochila [link] [comments]
submitted by /u/Informal-Trouble2183 [link] [comments]
DISCLAIMER: No Ai was prompted in the creation of this post. Today I had an experience that complely blew my mind, I just had to write…
submitted by /u/FlowCritikal [link] [comments]
This seems like a pretty solid improvement, and should make the more extreme quant setups viable on AMD cards. submitted by /u/Betadoggo_ [link] [comments]
https://github.com/woct0rdho/transformers5-qwen3.5-recipe It's time for GGUF to replace bitsandbytes as the base model format for low-VRAM LoRA training. It's actively supporting new model types such as MoE,…
I know DavidAU gets a bad rap, and rightfully so. I've tried some of his fine tunes in the past and they have been... Interesting. I…
It can only run the 1_0 quant in LM Studio. Only 3.8GB and runs on low-end hardware. Even a phone! https://prismml.com/news/bonsai-27b https://huggingface.co/prism-ml/Bonsai-27B-gguf 1 upvote submitted by…
I got my ML model training pipeline to go from 36 steps/minute to 47 steps/minute by optimizing how they're stored on disk. When training larger ML…