Show HN: Qwen Scribe – local transcription and dictation for Apple Silicon
Article URL: https://github.com/VladUZH/qwen-scribe Comments URL: https://news.ycombinator.com/item?id=49098260 Points: 8 # Comments: 0
Article URL: https://github.com/VladUZH/qwen-scribe Comments URL: https://news.ycombinator.com/item?id=49098260 Points: 8 # Comments: 0
I was looking for a strong coding model and a strong general model, both should be 120b or under. After weeks of researching, qwen3.6 27b(general) and…
RT Artificial AnalysisAlibaba has released Qwen Audio 3.0 Realtime, with the Plus variant debuting as the new #1 model on the Artificial Analysis Speech to Speech…
Hi everyone! We’ve just released a major update to the leaderboard! We are expanding beyond Python with a new multilingual slice featuring real-world software engineering tasks…
Since launching Qwen3.8-Max-Preview, we've received valuable feedback from developers!To help users better explore Qwen3.8's agentic capabilities, we're officially launching the #QwenGrowthPlan today! 🚀We invite you to:-…
I'm not quite following the news. Is it true? When will the Qwen versions of this size be released? submitted by /u/dai_app [link] [comments]
From the description: "Reasoning-Medical-27B is designed for universal advanced medical reasoning in professional medicine, medical genetics, college biology/medicine, and clinical knowledge. The model was fine-tuned on…
Most quantization works like this: pick a bit depth, apply it everywhere, maybe let imatrix take a rough guess at what matters, ship it. Most don't…
submitted by /u/fulgencio_batista [link] [comments]
I’ve been doing some benchmarking on my mini-PC setup (AMD Ryzen 7 6800H) to see how it handles the new Gemma 4 and Qwen 3.6 MoE…
I just managed to get it running on windows and this thing is fucking insane. I get around 550-720t/s depending on task at hand. Previously to…
submitted by /u/pmigdal [link] [comments]
I finished the speed leg of my spec-decode benchmarking for Qwen3.6-27B, main algorithms across quants. Overall: the heavier the quant, the more spec-decode buys you (10…
Forked SGLang, wrote TeilLang FlashAttention for V100, used open-source marlin-v100, ungated flashinfer for sm70, made Dflash work for Qwen3.5/3.6 models, added Laguna S2.1 support, tried to…
Test Prompts: 1.1. Algorithm & Logic (10 pts): "Write a function in Python that finds the contiguous subarray with the largest sum (Kadane's algorithm). Include time…
arXiv:2607.21774v1 Announce Type: new Abstract: Large language models may infer demographic attributes from subtle linguistic cues even when those attributes are not explicitly stated. This pilot…
decided to mess around with it to see how the world model training affects it, found a system prompt that massively improves reasoning: predict your own…
Instead of 2T+ models, continuing to release highly capable small to medium size LLMs would really help to keep this community vibrant. Hardly anyone can even…
Qwen3.6-27B: SFT vs continued pre-training vs RL? I’m interested in adapting Qwen3.6-27B, but I’m increasingly unsure whether conventional SFT/LoRA is the best route if the goal…
Last week someone here said ThinkingCap and Fable Fusion "really do beat the OG" for agentic work, so I ran it: 6 self-grading tasks, 5 reps,…
It came out 3 days ago just wondering if anyone's tried it yet? submitted by /u/MundanePercentage674 [link] [comments]
I am writing a fairly complex C++ windows MFC application. I have a few 3090s and can run F16 Qwen3.6-27B with 256K context and MTP. The…
Benchmarks using single system running triple GPU with 31GB Vram combined. NVIDIA GeForce GTX 1080 Ti 11GB (NVIDIA) NVIDIA P102-100 10GB (NVIDIA) - first instance (distant…
I understand that the laguna model is either still buggy or potentially benchmaxxed. So I’d like to know for people who really tested, are DS flash…
Qwen is Alibaba's open-weight model family — Qwen3 (text), Qwen3-VL (vision), Qwen3-Coder, QwQ (reasoning). Qwen3 sits near the top of every open-weight benchmark in 2026 and ships under a permissive licence.
Owner: Alibaba. We have 350 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is qwen.ai.
Related text models: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, DeepSeek.