Luth-2: New State-of-the-Art French Small Language Models
Hey everyone, Today we release Luth-2-0.8B and Luth2-2-2B, two non-reasoning models that set a new state of the art for French across a wide variety of…
Hey everyone, Today we release Luth-2-0.8B and Luth2-2-2B, two non-reasoning models that set a new state of the art for French across a wide variety of…
I'm very happy with some Linux admin tasks I'm throwing at a locally running DeepSeek. My request was simple, check why 'samples' folder is taking more…
I am a Macbook Pro M4 user with the 16GB of unified ram. The best model I have been able to run on LM Studio is…
Seems like we're limited to qwen finetuned MoEs for now. Looking at the current landscape - focus seems to be on dense models (muse glimmer 30b,…
Available b10356 onwards. Overview ROCm 7.14 is the first production release using TheRock build system. It can be installed using multi-arch deliverables from wheels, debs, rpms,…
Heeeey all! I just completed some fun tests with Muse Glimmer, I thought I'd let you know. In fact, the summary below was written by Muse…
Confirmed by the official Qwen account. submitted by /u/Bestlife73 [link] [comments]
submitted by /u/fallingdowndizzyvr [link] [comments]
I wanted to find out whether a huge text-only MoE could be given basic vision without retraining the language model itself. The short answer is yes.…
A few things right off the bat: it reasons very efficiently. Like Grok 4.5 levels of efficient thinking it quantizes very well. My first few tests…
r/LocalLLaMA is one of 177 primary AI sources we aggregate. 2001 stories from this source have been indexed. Domain: www.reddit.com. All posts here link straight to the original — we don't republish content, we point readers at it.
See the full source catalogue or browse by model.