Will we have a 27B model with Fable capabilities in 5 months? History says yes
If history is any indication, open-source models in the 27B dense range should have caught up to what the US government banned two weeks ago because…
If history is any indication, open-source models in the 27B dense range should have caught up to what the US government banned two weeks ago because…
I’ve always had this idea but can’t prove it. I think Anthropic and OpenAI don’t really have any secret sauce, their moat is just scale. Rumor…
I have always said we should build huge rigs, figure out how to run the largest models for moments like this. If Kimi K3 is near…
Hey guys, Last week I posted my DFlash benchmarks here (4.44x at 36K context) https://www.reddit.com/r/LocalLLaMA/comments/1uq0h4o/i_tested_freshly_merged_dflash_in_llamacpp_on/ . One of the comments form u/exact_constraint mention there are also…
With all of the attention Kimi K3 is getting at the moment and likely for the coming weeks, would we be seeing Dario Amodei continue to…
Unbelievable to see kimi k3 beat frontier models that were 'too dangerous' for public use. submitted by /u/Gohab2001 [link] [comments]
The gap is no longer measured in months. Open-weight models are catching up in real time, and Kimi K3 is already performing near Fable level. At…
submitted by /u/coder543 [link] [comments]
submitted by /u/MagicZhang [link] [comments]
submitted by /u/fallingdowndizzyvr [link] [comments]