DeepSeek v4 Flash 0731 locally on CPU
After seeing the benchmark results for the full release of DS v4 Flash 0731, I replaced my 2 x 16GB DDR4 ram sticks with 2 x…
After seeing the benchmark results for the full release of DS v4 Flash 0731, I replaced my 2 x 16GB DDR4 ram sticks with 2 x…
There are so many posts where people complaining about high prices and asking for solution <= 1000 EUR. So, there is one solution to consider: PC/mini…
https://tencent-hunyuan.github.io/Hunyuan3D-WorldClaw/ Looks impressive from that site, hopefully they open weight this so we can all play with it. submitted by /u/Uncle___Marty [link] [comments]
The gods have blessed me. Card came in. brand new. I couldn’t believe it. submitted by /u/Street-Buyer-2428 [link] [comments]
Available context length with and without the patch: Model: QWEN 27B ROCm stock patched Vulkan stock patched IQ4_XS Pure, single 16GB GPU 19.456 76.032 68,352 78,592…
The past week I've been running DSv4 inference on my laptop by keeping everything RAM-resident except the MXFP4-experts (since expert pool is ~147GB and won't fit)…
Hi, After some tries, it seems that for local models (27B+) the best way to have reliable outputs is to add a little more code and…
The past week I've been running DSv4 inference on my laptop by keeping everything RAM-resident except the MXFP4-experts (since expert pool is ~147GB and won't fit)…
Disclosure: I’m the author of Ante. DeepSeek recently reported an 82.7% score on Terminal-Bench 2.1 for DeepSeek V4 Flash 0731. Its evaluation used “DeepSeek Harness minimal…
What Local Embedding + Reranking Models are you guys running for RAG? I went down this rabbit hole because I wanted a Embedding Model + Reranker…