Any word on Qwen 3.7 9B? (Also looking for 9B-class alternatives to Qwen 3.5)
Given that Alibaba went proprietary/API-only for the Qwen 3.7 Max and Plus launches back in May, do we have any rumors or roadmap for a local…
Every primary-source story across every tracked model. Filter by clicking a chip.
Given that Alibaba went proprietary/API-only for the Qwen 3.7 Max and Plus launches back in May, do we have any rumors or roadmap for a local…
RT Maksym Andriushchenko @ ICML 🇰🇷i highly doubt that GLM-5.2 was benchmaxxed on PostTrainBench or heavily distilled from Claude models. anyone can inspect the traces (https://posttrainbench.com/traces/):-…
Article URL: https://github.com/Trystan-SA/claude-design-system-prompt Comments URL: https://news.ycombinator.com/item?id=48792399 Points: 23 # Comments: 1
I made a 10MB LoRA adapter for Qwen3.5-4B plus a small orchestration layer. It decides, per query, whether to answer directly, search the web, or retrieve…
As someone who is not affiliated with any of the big tech companies, I find it particularly difficult to have the confidence or enthusiasm to approach…
Kling AI NEXTGEN Awards Ceremony | PreviewStay tuned!Date: July 7, 2026 | 14:00–18:00Venue: Seoul Film Center
Junyang Lin, the former technical lead of Alibaba's Qwen, walked through the model family in a talk "towards a generalist model / agent," then expanded it…
Not really a tutorial, but more of sharing my attempts at getting higher contexts on Q8 of Qwen3.6-27 with 32GB VRAM. Disclaimer: Not in-depth research. Crowd…
Somewhat humbling to have Claude Fable do a final review of some software that you're about to release and have it then find (and fix) FIVE…
I wrote about the sqlite-utils 4.0rc1 release a couple of weeks ago. Since we only have Claude Fable on our Max subscriptions for a few more…
Article URL: https://github.com/openai/codex/issues/30364 Comments URL: https://news.ycombinator.com/item?id=48789428 Points: 9 # Comments: 1
submitted by /u/johnnyApplePRNG [link] [comments]
The open-source tool pxpipe converts long text prompts for Claude Code into compact PNGs, exploiting the fact that Anthropic charges for images by pixel size, not…
As part of an ongoing legal dispute with three Hollywood studios, Midjourney is seeking to compel those studios to reveal how they use AI themselves.
I've mentioned this kernel project I was working on in a few posts and figured I would just open the project code for anyone curious: MLX…
Check it out: https://github.com/fairydreaming/llama.cpp/tree/dsv4 They are PRs #25247, #25303 (mine) and #25202 (from am17an) but I omitted some padding changes from the last one that I…
Alibaba has reportedly classified Claude Code as high-risk software.
Anthropic released Claude Science in beta on June 30, 2026. The app runs on existing Claude models. A coordinating agent delegates to domain specialists, a reviewer…
Mistral AI, which offers some open source AI models, has raised significant funding since its creation in 2023, with the ambition to “put frontier AI in…
Threw together a benchmark suite (quest completion, scene endings, item/time tracking, character detection, storytelling, drafting) and ran it across 8 models people talk about a lot…
Today I saw this in StepFun’s blog for their Step 3.7 Flash model. Running the model with CC performed better results vs Hermes. Curious why? submitted…
Spent a while tuning llama.cpp for Qwen3.6 27B on a 9800X3D / 64GB / 5090 box and wanted to share the real distribution instead of just…
how do you like this Fabulism?It's growing on me, way better than Opus's lame PUSHBACK frictions. this is an AI built for stateful agents improving via…
A lot of people seem to be confused or mystified about this so figured I'd spell it out. I played around with RYS and realized that…
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.