DeepSeek V4 Flash unsloth quants are out!
4-bit - 155gb 8-bit - 162gb "Smaller ones are coming" submitted by /u/RunawayPeeko [link] [comments]
4-bit - 155gb 8-bit - 162gb "Smaller ones are coming" submitted by /u/RunawayPeeko [link] [comments]
Source: https://x.com/deepseek_ai/status/2083084415157022911 & https://deepswe.datacurve.ai/ just combined data view. DeepSeek claims, not verified by DeepSWE yet. submitted by /u/sdexca [link] [comments]
Tell me this is not funny - "OH MY GOD. I THINK I FINALLY SEE IT!!! The black pixels are at the QUAD CENTERS because of…
Deepseek's budget model V4 Flash gets a major boost with the "0731" update, jumping ten points to 50 on the Artificial Analysis Intelligence Index. That puts…
submitted by /u/Antique_Archer_7110 [link] [comments]
While waiting for some of the quants to drop, I load the API with $50 and ran it on SlopCodeBench Just vibe reading the results it…
submitted by /u/BlackBeardAI [link] [comments]
Didn't see anything about it in their announcement submitted by /u/Eyelbee [link] [comments]
Remember when R1 had Llama and Qwen distills? Can we expect those for v4? submitted by /u/Aggravating-Push-207 [link] [comments]
Huawei trains, basically, a DeepSeek V3.5, except for minor mHC/ModAttn modifications. Whale Paradigm reigns more supreme than ever.But more importantly: second large model on Ascends.There is…
RT kabikabideepseek flash确实强啊,来感受一下这个长程能力。agentic能力现在强太多了。甚至他自己发现了harness里面 agent swarm 的 tool call,并且做好了拆分和安排。这个是提示词里面没有的,它自己发现并且组合使用subagent swarm模式的。之前测的时候,flash太蠢了。别说调用复杂工具链进行切分和安排。就是tool call都容易出问题。截图是Deepseek + maka,https://github.com/maka-agent
RT 刘叉|增长运营Re @lifesinger 奥特曼刚刚发起价格战,就被梁圣阻击。将DeepSeek V4 Flash标到图中,才看得出这性价比真的够刚https://x.com/haoliucha/status/2083121333916102820?s=20刘叉|增长运营: OpenAI试图重返AI王座,A畜暂且装死,DeepSeek相应迅速发起了反阻击战。DeepSeek-V4-Flash的性价比到底有妖孽,你看👇下图就知道了。
«DeepSeek has missed the boat on agents»Uh huhWatch themWeihao Zeng: Try our model on general agent tasks —making intelligence more accessible to everyone 🥳🥳🥳
RT Lincoln 🇿🇦All takes from last night about China being cooked won't age well.Again, we don't say thank you to DeepSeek enough.Only ones pushing for lower…
RT ZiwenCodex can now run Deepseek-v4- flash!There's a catch though. Deepseek's official setup switches your entire codex over to them, so your GPT models stop showing…
Amazing innovation in chart crimes here @ArtificialAnlys I get that Sama needs his victory lap, many fools asked me "how will DeepSeek recover after Luna price…
Article URL: https://artificialanalysis.ai/models/deepseek-v4-flash-ga Comments URL: https://news.ycombinator.com/item?id=49120299 Points: 10 # Comments: 2
submitted by /u/MagicZhang [link] [comments]
RT Max For AI声明:本人从未在互联网上诋毁过深度求索,并没有称梁文锋为梁白开。本人实际上是深度求索一年半年老粉,经过长时间的思考,我发现DeepSeek才是中国AI的希望,有时候做出决定很难,经过许多个日夜的思考,我决定加入DeepSeek粉丝团。至于Kimi,GLM和Minimax,祝他们的开源模型一切顺利。本人从未诋毁过深度求索,也从未称呼梁总为牢梁,望周知🥹
> these scores> our upcoming DeepSeek Harness (minimal mode)Who was saying that DeepSeek doesn't understand the importance of agents? Huh? Huh? I don't hear you!DeepSeek: 🚀…
RT DeepSeek🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta!🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out…
> when I launch Ainiux (soon OSS), then just run: ainiux deepseek -m "deepseek-v4-flash" --security-review on your code base. Grab a coffee (small code base) or…
Deepseek's new model V4 Flash 0731 is much better, I (Claude lol) did a bit of linear regression with a leave one out style verification to…
DeepSeek V4 Flash: Preview → 2026-07-31 Benchmark Preview 0731 Δ Terminal Bench* 56.9 82.7 +25.8 Toolathlon 51.8 70.3 +18.5 NL2Repo — 54.2 new Cybergym — 76.7…
DeepSeek shocked the field in late 2024 with DeepSeek-V3, then again with R1, a reasoning model trained at a fraction of the budget Western labs spend. Current lines: V3.2-Exp, R1, V3. The most-talked-about Chinese AI lab of the cycle.
Owner: DeepSeek. We have 280 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is www.deepseek.com.
Related text models: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen.