DeepSeek V4 Flash unsloth quants are out!
4-bit - 155gb 8-bit - 162gb "Smaller ones are coming" submitted by /u/RunawayPeeko [link] [comments]
4-bit - 155gb 8-bit - 162gb "Smaller ones are coming" submitted by /u/RunawayPeeko [link] [comments]
Source: https://x.com/deepseek_ai/status/2083084415157022911 & https://deepswe.datacurve.ai/ just combined data view. DeepSeek claims, not verified by DeepSWE yet. submitted by /u/sdexca [link] [comments]
A100 with 40gb VRAM: 162GB Q8_K_XL ~16.1 tok/s generation Only 15.8GB of 40GB VRAM used with all experts on CPU NOTE just tested coding on linux…
Tell me this is not funny - "OH MY GOD. I THINK I FINALLY SEE IT!!! The black pixels are at the QUAD CENTERS because of…
First of all, I want to apologize if it's off-topic or in the wrong format. Having tried Deepseek Flash with reasoning high on a conceptually difficult…
Been working hard for the past month to bring to the community some interesting curios, so for starters we have LongCat-Flash-Lite Uncensored Heretic with MTPs which…
The GGUFs are dropping! https://huggingface.co/unsloth/DeepSeek-V4-Flash-0731-GGUF submitted by /u/Omnimum [link] [comments]
submitted by /u/Antique_Archer_7110 [link] [comments]
submitted by /u/pmigdal [link] [comments]
While waiting for some of the quants to drop, I load the API with $50 and ran it on SlopCodeBench Just vibe reading the results it…