Qwen says next week 3.8 will be open weights
Should we ignore the hype cycle or play along? submitted by /u/Terminator857 [link] [comments]
Should we ignore the hype cycle or play along? submitted by /u/Terminator857 [link] [comments]
Spent way too much time with V4-Flash-0731 this weekend and wanted to share my vibes as briefly as possible. I sent it through a bit of…
Id figured since they first emailed people about api price changes coming mid july then delayed the v4 flash release to late july, I wonder if…
been thinking about how the desktop 70-class has sat at 12GB for two generations now, 4070, 4070 super, 5070, all 12GB. the 1070 gave you 8GB…
I tested them by sending the 6 documents, each meant to represent a different document type, through my own webapp and comparing every output against the…
Now we can run Qwen3-Next at “full speed” :) Do you still remember this model? submitted by /u/jacek2023 [link] [comments]
Hi everyone, if you have been reading the tech reports of Kimi, DS, Qwen and GLM, you will realize how much on policy distillation and GRPO…
is all about post training? idk much about that spesific thing, whats matter most to achieve more performance, optimized results i wonder submitted by /u/Initial-Carry1038 [link]…
It is so good! I don't know why there aren't more people talking about it. Fewer tokens, faster and more accurate than Qwen 3.6 35b a3b.…
https://github.com/zai-org/z-ai-sdk-java/commits/glm-5.3 submitted by /u/Few_Painter_5588 [link] [comments]