Can’t wait to see Qwen3.8-27B
Qwen announced Qwen3.8 a few hours ago, and it looks like we’re getting a new 27B model! Really excited to try this one locally. I’ll also…
Qwen announced Qwen3.8 a few hours ago, and it looks like we’re getting a new 27B model! Really excited to try this one locally. I’ll also…
submitted by /u/Hannibalj2ca [link] [comments]
Am looking forward to this! Open weight has come near frontier for 5x less the cost per/M tokens on task completion Open weight ranking Kimi k3…
MiniMax H3 is a general-purpose, omni-modal generative system. It supports unified understanding of multimodal contexts composed of text, images, video, and audio, and can generate video…
I like the outputs from this model, but DAMN does it over think. Has anyone found a robust fix for this that isn't just capping output…
https://preview.redd.it/gy0tgokdl2hh1.png?width=540&format=png&auto=webp&s=7db9e034613a915cb33d378b99ad72c31c7cc18f source: https://x.com/Alibaba_Qwen/status/2084100707423289643 submitted by /u/TKGaming_11 [link] [comments]
https://qwen.ai/blog?id=qwen3.8 submitted by /u/CounterReady4774 [link] [comments]
2.4 T parameter model. open weights coming soon! https://x.com/Alibaba_Qwen/status/2084093402967396594?s=20 submitted by /u/Mobile-Pumpkin7944 [link] [comments]
Kindly Benchmark Higher Quants of DeepSeek-v4-flash Against Qwen-3.6-27B Q8! I am running the UD-Q2_K_M of the model locally, though I can run Qwen3.6-27B_Q8_K_XL at around 70t/s…
WASTE is an embeddable inference engine written in C, with no third-party runtime dependencies. It keeps the model trunk in memory, streams selected experts directly from…