GLM 5.3 Spotted
https://github.com/zai-org/z-ai-sdk-java/commits/glm-5.3 submitted by /u/Few_Painter_5588 [link] [comments]
https://github.com/zai-org/z-ai-sdk-java/commits/glm-5.3 submitted by /u/Few_Painter_5588 [link] [comments]
AI9Stars has released G9v3-39A5B an open weights language model designed to deliver even stronger reasoning capabilities than ai9stars/G9v3-3B with its 39B and 5 active experts. It…
can configure any which way, tried a community deepseek nvfp4 but got a lot of issues. DeepSeek mxfp4 worked better but only with dspark off. looking…
Qweb 3.8-Max is out. #2 in Vision Arena, only behind Claude Fable 5 Max. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B…
Hey there! So today we're releasing SupraBrain-50M, a hybrid language model that combines Gated DeltaNet linear recurrence with Sliding-Window Attention and Surprise-Gated update mechanisms to deliver…
https://preview.redd.it/3zcvpbds14hh1.png?width=1911&format=png&auto=webp&s=a79aafb71eeca97638da93d2591902631e897fd5 I tried a test similar to the recent model-quant comparisons, but this time I focused only on: DeepSeek-V4-Flash-0731-IQ2_XS-Experts-Q8_0 model link bullerwins/DeepSeek-V4-Flash-0731-GGUF · Hugging Face
Google Cloud published OKF (Open Knowledge Format) on June 12th — a spec for storing curated knowledge as a directory of markdown files with YAML frontmatter.…
Super excited about this release for the new 27B. Who else is with me. Only 17GB VRAM needed 😍😍 submitted by /u/quantier [link] [comments]
Qwen announced Qwen3.8 a few hours ago, and it looks like we’re getting a new 27B model! Really excited to try this one locally. I’ll also…
submitted by /u/Hannibalj2ca [link] [comments]