Facing US export controls, China's DeepSeek plans to make its own chips
It's early, but the plan is to reduce dependency on Nvidia and Huawei.
It's early, but the plan is to reduce dependency on Nvidia and Huawei.
So DeepSeek got investment and is back to the chip game; glad that this is official. Wenfeng is an electrical engineer to begin with, and his…
Chinese startup Deepseek is building its own AI chip, Reuters reports. The article Deepseek is designing its own AI chip appeared first on The Decoder.
For quick look --> https://x.com/i/status/2059618247553745204 Detailed --> https://mimo.xiaomi.com/blog/mimo-v2-5-inference I hope in future we get fable lvl ai at the cost of current DSV4. Thats far more…
reminder that within 2 weeks, I expect a trifecta of DeepSeek Moments:- literally DeepSeek V4- SuperPod 950 (compute)- LandSpace ZQ-3 landingif they all go well, you'll…
> serve at DeepSeek-API-killing prices. That loop is why xAI runs Grok on SGLang and third parties> beat DeepSeek's own API by 5x on cost.what is…
Bearish for DeepSeek's roleplaying strategyTechmeme: ByteDance's Doubao and Alibaba's Qwen will disable humanlike and user-created agents before July 15, as China's anthropomorphic AI interaction rules take…
Hi. I don't have a GPU. So my biggest "local LLM" experience has been running ~26B models with single-digits tps values. However, the "serving economy" of…
Good if true, nice to see news of someone joining DeepSeek rather than the other way around once in a while. But… very many strong people…
Follow-up to my earlier posts: Should I sell my Mac Studio? https://www.reddit.com/r/MacStudio/s/GK7QP8Lg87 Kimi benchmark: https://www.reddit.com/r/LocalLLaMA/s/ujBsYLYmpd Short version: my Mac Studio was sitting mostly idle, and from…
Check it out: https://github.com/fairydreaming/llama.cpp/tree/dsv4 They are PRs #25247, #25303 (mine) and #25202 (from am17an) but I omitted some padding changes from the last one that I…
> Blackwell's increased memory bandwidth and communication bandwidth, plus DeepSeek's megamoe operator (though it seems to only support fp4?), make achieving 300tps on GLM 5.2 not…
> deepseek-v4-pro-202606> deepseek-v4-flash-202605It seems that V4 Release Version is almost here, and so is DeepSeek's coding plan. Except… it's still per-token?Saint Liang and @victor207755822 are true…
the second funniest thing is that Russian boomers will probably still trust Claude over DeepSeek even for their war-critical infra. Good American Quality!(but the funniest thing…
Here is the results of optimizing it for my setup: Benchmark results of the optimisation showing TG T/S from 22.7 to 21.3, and PP T/S from…
You may remember my earlier posts about DeepSeek V4 Pro at home. Today I checked the performance in my llama.cpp branch that contains various fixes and…
Hi folks. I found this video explaining latest DSpark breakthrough from Deepseek. Seems like a huge change coming. https://www.youtube.com/watch?v=J0D7qV3nl7w submitted by /u/BringTea_666 [link] [comments]
This is a follow-up to post about which local models stay fast deep into long context and I learned a lot from people here. I kept…
Should have just been copying DeepSeek all alongor idk, GLMAt Meta's scale, would be enoughbut they somehow never grew up to the point of accepting thisAndrew…
Wanted to try running DeepSeek V4 Flash locally but found it asking for absurd amounts of VRAM at higher context lengths (~256GB at 1M). Turned out…
RT Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)Re @meowbooksj the design space of sparsity is barely exploredLike, why don't we use MLP routers in MoEs? Because le…
Very interesting thesisDSA and similar inventions will certainly influence hardware design. DeepSeek isn't content to hope they'll win the hardware lottery, they'll choose the winning tickets.GDP:…
The xiaoren are not giving up!DeepSeek sees itself as a company that is building AGI. What has changed was the scale and the maturity of the…
Deepseek Flash V4 at IQ2 or Qwen 3.6 27B Q5KM ? Any tests or benchmarks ? Wondering which one would be better at speed / coding…
DeepSeek shocked the field in late 2024 with DeepSeek-V3, then again with R1, a reasoning model trained at a fraction of the budget Western labs spend. Current lines: V3.2-Exp, R1, V3. The most-talked-about Chinese AI lab of the cycle.
Owner: DeepSeek. We have 280 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is www.deepseek.com.
Related text models: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen.