I RL-trained Qwen3.6-35B-A3B to RL-train small task-specific Qwen models. Fully open source! 🤓
👋 Training my first RL model last year was super fun, now I've RL-trained a model that RL-trains other models... wild times! The agent gets a…
👋 Training my first RL model last year was super fun, now I've RL-trained a model that RL-trains other models... wild times! The agent gets a…
We tried to gather as many useable beneficial fable and opus 4.8 examples as possible. Fully uncensored (huihui) and we think the numbers speak for itself.…
RT Benjamin MarieThinkingCap-Qwen3.6-27B is really good.Much faster than the original Qwen3.6-27B thanks to shorter thinking, while preserving accuracy.Also one of the best evaluation I have seen…
Here model: https://huggingface.co/LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V2-GGUF Settings: temperature=0.6, top_p=0.95, top_k=20, min_p=0, presence_penalty=disabled, repeat_penalty=disabled, seed=42, thinking=enabled System Prompt: You are Qwen, a large language model created by Tongyi Lab team…
By the end of the year we should have:GPT 6Fable 5.5Gemini 3.5 ProGrok 5Spark 2Kimi 3Minimax M3.5GLM 6DeepSeek v4.5Mistral 4Qwen 4MiMo 3Never in the history of…
I tested Unsloth's Qwen3.6-35B-A3B-MTP-GGUF:UD-Q4_K_XL with simple Pi Coding Agent. No skill or extensions used. My laptop have 32GB RAM and 8GB VRAM. Without context it runs…
Hi everyone, I've spent about four days now trying to find the eight configuration for running Qwen 27b in production using VLLM but have been getting…
Hey everyone. Trying to figure out the best setup for my hardware. I've got a 64GB DDR5 RAM laptop and A 20GB 7900xt egpu. Usecase is…
The startup, PrismML, said it has shrunk down Qwen 3.6, an open-source large language model developed by Chinese internet giant Alibaba, to run on an iPhone…
Hey everyone, I recently switched from DS4 Flash to Qwen3.5-122B on my M3 Ultra Mac Studio for long-context agentic coding. While the model fit better, I…
Hey friends, I started a project i think we can all benefit from. I rented a few h200's and am fine tuning a hui Qwen 35b…
As many of you know t/s is super important. It's how fast your stuff gets done. I create via open code benchtest and run it. Thanks…
Let's start a discussion about what can be done to make local models more reliable. I've been using Qwen3.6-27B a lot lately, and have noticed the…
I recently posted some posts with VLLM showing issues with TTFT and concurrency with 4x 5060 ti's. Wanted to share this benchmark to provide what worked…
Hey all, here are two new high performance qwen3.5 gguf sets I created using a new state of the art technique for optimizing mixed precision called…
I got a new mini-pc for a homelab server recently and thought I'd tinker around with some LLM options on there. As it doesn't have a…
how much vram do you need and what model do you think is the next major upgrade from the good old qwen 3.6 27b as of…
Article URL: https://mrzk.io/posts/qmlx-maximising-ai-psychosis-minmaxing-mac-studio/ Comments URL: https://news.ycombinator.com/item?id=48876619 Points: 5 # Comments: 0
WEIRD DISCLAIMER: none of this was written by an LLM until you get to the Github repo/site, which was obviously assembled by your friend and mine,…
Which settings would suffice to work with it ? submitted by /u/soyalemujica [link] [comments]
How far can i stretch the context window with Qwen 3.6 27B (using Q8_0) before it gets too unreliable? I am at 100k right now and…
Hey guys, we’ve built fastest speculative decoding for Qwen at least. To run in sglang you can use our fork Hf link Would appreciate your feedback…
Experimented with some custom CUDA and C++ code that can now run a Qwen3-30B-A3B at 50-54 tok/s at float 8 on an RTX 5060 Ti with…
So, umhh, I am working on an agentic coding platform, and I need to make qwen3.5 and gemma4 models out of controlled reasoning chains. For example,…
Qwen is Alibaba's open-weight model family — Qwen3 (text), Qwen3-VL (vision), Qwen3-Coder, QwQ (reasoning). Qwen3 sits near the top of every open-weight benchmark in 2026 and ships under a permissive licence.
Owner: Alibaba. We have 350 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is qwen.ai.
Related text models: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, DeepSeek.