DeepSeek-V4-Flash-0731 Open weight!
deepseek-ai/DeepSeek-V4-Flash-0731 · Hugging Face submitted by /u/shing3232 [link] [comments]
deepseek-ai/DeepSeek-V4-Flash-0731 · Hugging Face submitted by /u/shing3232 [link] [comments]
Was playing around with TurboFieldfare, a Mac engine that runs Gemma 4 26B in ~2 GB by streaming MoE experts off SSD instead of loading them.…
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731 submitted by /u/cgs019283 [link] [comments]
Article URL: https://andre.arko.net/2026/07/30/ruby-centrals-destructive-legacy/ Comments URL: https://news.ycombinator.com/item?id=49122105 Points: 9 # Comments: 2
One day before OpenAI’s HF incident disclosure, OpenAI disclosed that it paused internal deployment of a long-horizon model after it circumvented its sandbox, then restored access…
Article URL: https://tasklet.ai/careers/customer-success-engineer Comments URL: https://news.ycombinator.com/item?id=49122034 Points: 0 # Comments: 0
Didn't see anything about it in their announcement submitted by /u/Eyelbee [link] [comments]
Article URL: https://hughhowey.com/the-end-of-an-era/ Comments URL: https://news.ycombinator.com/item?id=49121980 Points: 64 # Comments: 44
I have a 5090 in my main PC and a 3090 in my last PC. I was planning on selling the older PC but now thinking…
Article URL: https://arxiv.org/abs/2607.27197 Comments URL: https://news.ycombinator.com/item?id=49121868 Points: 34 # Comments: 17
Remember when R1 had Llama and Qwen distills? Can we expect those for v4? submitted by /u/Aggravating-Push-207 [link] [comments]
Okay, hear me out. Why is it that every time a new model comes out, all the tests I see are "make a car game," "make…
I keep seeing the "use a big model via API as the architect, run local small/mid models as workers" pattern recommended for people with modest local…
submitted by /u/SnooBunnies8392 [link] [comments]
submitted by /u/InternationalGap3698 [link] [comments]
Article URL: https://artificialanalysis.ai/models/deepseek-v4-flash-ga Comments URL: https://news.ycombinator.com/item?id=49120299 Points: 10 # Comments: 2
submitted by /u/MagicZhang [link] [comments]
Article URL: https://www.bbc.com/news/articles/cn0nqv05g0do Comments URL: https://news.ycombinator.com/item?id=49120120 Points: 6 # Comments: 3
Article URL: https://blog.google/security/chrome-stronger-with-every-update/ Comments URL: https://news.ycombinator.com/item?id=49120097 Points: 17 # Comments: 14
Beats GLM 5.2, and is the same cost as the previous one. submitted by /u/Potential_Top_4669 [link] [comments]
openPangu-2.0-Pro is an MoE model trained on Ascend. The model has 505B total parameters and 18B activated parameters. Its context length is 512k. The total pretraining…
Deepseek's new model V4 Flash 0731 is much better, I (Claude lol) did a bit of linear regression with a leave one out style verification to…
DeepSeek V4 Flash: Preview → 2026-07-31 Benchmark Preview 0731 Δ Terminal Bench* 56.9 82.7 +25.8 Toolathlon 51.8 70.3 +18.5 NL2Repo — 54.2 new Cybergym — 76.7…
Article URL: https://api-docs.deepseek.com/updates/ Comments URL: https://news.ycombinator.com/item?id=49119559 Points: 43 # Comments: 17