New DeepSeek V4 Flash 0731 vs ChatGPT Luna comparison
submitted by /u/perelmanych [link] [comments]
submitted by /u/perelmanych [link] [comments]
The model refuses to load into VRAM and uses only RAM. What can be an issue? Q2_K_XL from Unsloth if that changes something. submitted by /u/esw123…
March 6th, 2026 the highest intelligence index score was 51 for frontier models. deepseek-ai/DeepSeek-V4-Flash-0731 that has an intelligence score of 50. If these benchmarks are accurate,…
I don't know much about llms aside from downloading them through a frontend and running them on my laptop or potato phone. Google released gemma 4,…
submitted by /u/galapag0 [link] [comments]
CPU: Threadripper 3970X RAM: 128GB DDR4 GPUs: 3x2080ti 11GB The current best parameters to run it: llama-server --model Qwen3.6-27B-Q5_K_S.gguf --n-gpu-layers 999 --split-mode tensor --flash-attn on --cache-type-k…
https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/ Even after the 90% cut, Luna is STILL about 3.3x more expensive than DS4F for the same intelligence bracket. From the chart, the obvious next…
sat on a mat. submitted by /u/Risen_from_ash [link] [comments]
I'm constantly bombarded by non-local LLMs in this sub but god forbid I post a local model meme. submitted by /u/fragment_me [link] [comments]
What speeds are everyone getting with deepseek v4 flash 0731? I’m getting~200 tps prompt processing / ~11 tps token gen, on 4x5060ti16gb with ddr4 3200 ram…