What’s real value of high reasoning?
Guys just I have simple question as I have seen majority of models does well with medium reasoning. Qwen 36 3.5 pull even well with no…
Guys just I have simple question as I have seen majority of models does well with medium reasoning. Qwen 36 3.5 pull even well with no…
https://huggingface.co/unsloth/DeepSeek-V4-Flash-0731-GGUF It's coming! 😍 submitted by /u/Treidge [link] [comments]
A lot of recent models are being announced with promised open weights, but the weights are either weeks away, or in some cases (looking at you…
GGUF is comming...I hope today UPDATE: It's there! and unsloth is comming https://huggingface.co/unsloth/DeepSeek-V4-Flash-0731-GGUF submitted by /u/mossy_troll_84 [link] [comments]
deepseek-ai/DeepSeek-V4-Flash-0731 · Hugging Face submitted by /u/shing3232 [link] [comments]
Was playing around with TurboFieldfare, a Mac engine that runs Gemma 4 26B in ~2 GB by streaming MoE experts off SSD instead of loading them.…
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731 submitted by /u/cgs019283 [link] [comments]
Didn't see anything about it in their announcement submitted by /u/Eyelbee [link] [comments]
I have a 5090 in my main PC and a 3090 in my last PC. I was planning on selling the older PC but now thinking…
Remember when R1 had Llama and Qwen distills? Can we expect those for v4? submitted by /u/Aggravating-Push-207 [link] [comments]