We could really use Qwen3.8 in 27B, 35B, 122B and 397B sizes
Instead of 2T+ models, continuing to release highly capable small to medium size LLMs would really help to keep this community vibrant. Hardly anyone can even…
Instead of 2T+ models, continuing to release highly capable small to medium size LLMs would really help to keep this community vibrant. Hardly anyone can even…
Running local models means every token counts — an 8K or 16K window fills up fast when you're stuffing graph context into prompts for RAG. I…
I got tired of my Ollama server sitting idle between chat experiments, so I pointed it at my Navidrome library and made it run a radio…
something big should happend for them to call "Note: All GGUFs generated before this change will need to be regenerated." llamacpp website submitted by /u/EconomySerious [link]…
I was refreshing my youtube and found out my favourite reviewer uploaded a battery test of 78 smartphones: https://youtu.be/MpgUFrsIWSQ the author said they started using robotic…
submitted by /u/Time_Reaper [link] [comments]
I am seeing more and more videos as posts about how OpenAI is in complete financial ruin, Anthropic isn't much better. Their expenses go with the…
Hey r/LocalLLaMA! :D I wanna share a really cool fully OSS thing I've been building that's only possible with local models: truly proactive AI! All your…
There is no quality or value to this post, however, I hope you might find humor in this broken output from my local Qwen TTS setup.…
Qwen3.6-27B: SFT vs continued pre-training vs RL? I’m interested in adapting Qwen3.6-27B, but I’m increasingly unsure whether conventional SFT/LoRA is the best route if the goal…