Qwen 3.8-27b coming this week
Confirmed by the official Qwen account. submitted by /u/Bestlife73 [link] [comments]
Confirmed by the official Qwen account. submitted by /u/Bestlife73 [link] [comments]
Just downloaded the model, UD-Q5_K_XL quant, asked it to generate a long story to test out reasoning and speed with dflash (super fast btw, ~ 90…
submitted by /u/ShadyShroomz [link] [comments]
There's an interactive chart and some extra data in the blog post if you're interested. There are plenty of KL-divergence benchmarks for GGUF models, but most…
To be clear Qwen 3.6 27B dates back to Apr 21, it's the same generation as V4-PreviewGemma is earlier in AprilSo most of this is just…
I've been comparing Cline / Kilo / Qwen Code lately since they all handle long-task state differently. Cline: has Focus Chain, a markdown file kept outside…
I also have 64 gb ddr4 ryzen 5600 Using llama.cpp Ubuntu distro Settings are as follows --n-gpu-layers 999 --n-cpu-moe 37 --no-mmap -ctk q8_0 -ctv q8_0 -fa…
TL;DR: I spent another 8 days following my last post making major improvements to the WinterMix method for MLX models. At 20k+ context this 59 GiB…
table bench https://huggingface.co/bartowski/endless-frontier_BigBang-v1-GGUF I'm downloading this model only because Bartowski converted it to .gguf, so it might be interesting. Doubts : The headline number is basically…
Available context length with and without the patch: Model: QWEN 27B ROCm stock patched Vulkan stock patched IQ4_XS Pure, single 16GB GPU 19.456 76.032 68,352 78,592…
Pasted the same HTML/JS code (330 lines) into Qwen 35B A3B and Gemma 26B A4B. Qwen: tokenized the input to 1609 tokens Gemma: tokenized the input…
submitted by /u/InternationalGap3698 [link] [comments]
Looking for V100 users to share your config and it's performance. GPU: Tesla V100 PCIE 32Gb Qwen3.6 27B Q4_K_M + Q8_0 MTP 128K context length Pi…
I've been tuning my new Radeon AI Pro R9700, and figured that this would be useful information for people who are trying to optimise their setups.…
I’ve been using 3.6 27B Q4, and that quant is fast on an M5. The code has been average, but consistently “good enough.” And, after a…
I compared Qwen 35B-A3B MoE against Qwen 27B dense on a series of local coding-maintenance tasks. On my R9700/llama.cpp setup, the MoE model generated about 3.9×…
I run the following on a 5090 and have been okay with its performance, it does most things somewhere 80-100 t/s, though that can slow down…
Alibaba plans to introduce revenue-sharing terms for some commercial users of its next Qwen open-weight AI model, Reuters reported, citing two people familiar with the company’s…
The cloud was always a cat. Qwen3.8-Max just saw it first. 😼☁️ Try it yourself!Ann Nguyen: Qwen 3.8 Max is actually impressiveSent a sky pic to…
Thanks for the thorough testing! With Qwen3.8-Max, everyone can observe the world in detail. 👀SkalskiP: 8.5 percentage points ahead of Gemini 3.5 Flashcongrats to @Alibaba_Qwen !
Hey peeps. I know you're tired of low quants giving hard to believe numbers. I'm quite skeptical too and from what I tried I'm often left…
First of all, my setup: Ryzen 9 5950x DDR4 3200Mhz 64gb (2x32) Dual 3090s, no NVLINK Runtime: llama.cpp Nvidia Drivers 610 Windows 11 25H2 Qwen 3.6…
submitted by /u/anderspitman [link] [comments]
Article URL: https://artificialanalysis.ai/?intelligence=agentic-index Comments URL: https://news.ycombinator.com/item?id=49200652 Points: 144 # Comments: 60
Qwen is Alibaba's open-weight model family — Qwen3 (text), Qwen3-VL (vision), Qwen3-Coder, QwQ (reasoning). Qwen3 sits near the top of every open-weight benchmark in 2026 and ships under a permissive licence.
Owner: Alibaba. We have 349 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is qwen.ai.
Related text models: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, DeepSeek.