Vision Support for Minimax-M3 has been merged into llama.cpp
submitted by /u/Time_Reaper [link] [comments]
submitted by /u/Time_Reaper [link] [comments]
I am seeing more and more videos as posts about how OpenAI is in complete financial ruin, Anthropic isn't much better. Their expenses go with the…
Hey r/LocalLLaMA! :D I wanna share a really cool fully OSS thing I've been building that's only possible with local models: truly proactive AI! All your…
There is no quality or value to this post, however, I hope you might find humor in this broken output from my local Qwen TTS setup.…
Qwen3.6-27B: SFT vs continued pre-training vs RL? I’m interested in adapting Qwen3.6-27B, but I’m increasingly unsure whether conventional SFT/LoRA is the best route if the goal…
curious how people here approach buying high-end/workstation cards (RTX 6000 Ada, 5000 Ada, etc) for local LLM work, do you actively watch pricing/timing on these specifically,…
submitted by /u/Unusual_Guidance2095 [link] [comments]
I ran DeepSeek V4 Flash through Claude Code, OpenCode and Pi on my own benchmark, and the quality came out basically the same across all three…
submitted by /u/RhubarbSimilar1683 [link] [comments]
Tell me, to get on 20k context and ingestion 44tks, generation 8tks is good numbers for 4x 8880 v4 cpus, 1tb 32channels ddr3 ram and 2x…