Bridging the Domain Gap: AI Race Coach built with Antigravity and Gemini
On May 23, 2026, fresh off the stage at Google I/O, our Google Developer Experts (GDEs) converged on...
Every primary-source story across every tracked model. Filter by clicking a chip.
On May 23, 2026, fresh off the stage at Google I/O, our Google Developer Experts (GDEs) converged on...
RT Luma | Dream LabThis looks like a $100K shoot. It started with one image.🌊 Consistent faces from reference images.🎥 Realistic water motion from camera movement.✨…
8K/GPU at 80 tps on B300 NVL72 with DynamoDeepSeek does 1K with DSpark on whatever it is they've got (and they do give me 80 tps…
surprising. Achiam was a very consistent OpenAI loyalist.Andrew Curran: Big news, Chief Futurist Joshua Achiam is leaving OpenAI. I wish him all the best in the…
for DeepSeek, being a "researcher" is not enoughHanchi Sun: @teortaxesTex If u read his zhihu posts he makes way too many consensus calls that were wrong,…
GLM 5.2 on Together AI is now #1 on Artificial Analysis for both output speed and latency.
MAI-1 has no independent benchmarks yet, but the ones they released suggest it is worse than Sonnet 4.6.Not sure it is going to be great as…
It's more of a showcase demo at the moment, but Qwen3TTS 0.6B runs usable fast in Q4. On my Galaxy S25 I get around 0.5 realtime…
Open-source models’ success isn’t coming at the expense of frontier labs. Instead, they each seem to capture two phases of the same life-cycle.
RT Arena.aiExciting news: Meta’s Muse Image just claimed #2 in the Image Arena!Muse Image from @AIatMeta now ranks second only to OpenAI's GPT Image 2, outperforming…
VisionBridge lets you give text-only LLMs vision. It's tiny OpenAI-compatible proxy that lets reasoning models (DeepSeek, Qwen, GLM…) see images by querying a separate vision model…
We have 8xB200 nodes and users keep asking us how to serve GLM-5.2 on them. Our engineering team went through everything published so far, and the…
submitted by /u/cuolong [link] [comments]
Microsoft is replacing AI models from OpenAI and Anthropic with its own MAI models in products like Excel and Outlook. Tens of thousands of queries per…
Cohere has released Transcribe Arabic, an open-source model for Arabic speech recognition that the company says outperforms Whisper and OmniASR on dialects, code-switching, and bilingual Arabic-English…
Starting Tuesday, Anthropic's Claude Cowork AI platform will be available on mobile and web for the first time. The expanded access is rolling out first to…
Anthropic is rolling out its AI agent Claude Cowork to mobile and web. Until now, the feature was limited to the desktop app. The agent keeps…
Hey guys, A month ago I posted my MTP benchmarks here (3.34x on Gemma 4). DFlash support just merged into llama.cpp (PR #22105), so I ran…
Lower is better - Quantization increases from right to left I recently made a post here about how I squeezed more context into a Q8 model…
!!Jukan says these guys are DeepSeek's silicon partner.Jukan @ ICML: @teortaxesTex verisilicon
With this update, users can start a task from their desk, get status updates on their phone, and pick up the finished output later — even…
It's early, but the plan is to reduce dependency on Nvidia and Huawei.
Claude’s desktop app is brilliant, but for our own daily work we kept wanting it to be less like a chat app and more like a…
RT HassanIntroducing AI Browser Games!Watch open & closed models build small browser games head to head.Open models like Kimi K2.7 were faster, cheaper, & produced games…
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.