Unsloth Quantization of Laguna S 2.1 Is Out
Various quantization now available, thanks Unsloth team ! submitted by /u/BoogerheadCult [link] [comments]
Various quantization now available, thanks Unsloth team ! submitted by /u/BoogerheadCult [link] [comments]
Properly done this time. Models as the GIFs are displayed: Qwen3.6-27B - 4bit Qwen3.6-MoE - 6bit Ornith-35B - 6bit Gemma-4-26B - 6bit Qwen3.6-MoE - 4bit HuiHui-Qwen3.6-MoE…
Nanbeige Lab released Nanbeige4.2-3B, and if the benchmark claims hold up, the numbers are pretty crazy for a model this small. It’s built on a "Looped…
Announcement tweet here. Direct link: Language Model Builder From the site: "Using the default settings, you’ll get a model that writes coherent, grammatical multi-paragraph text in…
So I was translating samples from Dolci-Think-SFT-7B and I thought it'd be an easy task, just deploy Gemma on vllm and write a quick translation prompt,…
Hey guys, I went ahead and installed a 100Gbe NIC card on both my MI50 machine and my P40 machine and loaded Nemotron Ultra IQ3_S across…
submitted by /u/Thrumpwart [link] [comments]
Despite local models getting significantly better, it seems that no one is trying to replicate the existing accomplishments of closed models. When it inevitably drops, would…
submitted by /u/Qwen30bEnjoyer [link] [comments]
https://x.com/ClementDelangue/status/2079670308156645882 PsyOP much when they claim defense with a Chinese model ? To be fair the attacking model made quite a ruckus when getting in so…