LWiAI Podcast #234 – Opus 4.6, GPT-5.3-Codex, Seedance 2.0, GLM-5
An action-packed episode!
Every primary-source story across every tracked model. Filter by clicking a chip.
An action-packed episode!
GLM-5 from Zhipu AI and Minimax M2.5 are now available in Windsurf with limited-time promotional pricing. Both models are included in Arena Mode's Frontier Arena and…
A crazy packed edition of Last Week in AI! Plus some small updates.
Ollama now supports subagents and web search in Claude Code.
Carlsen will bring his iconic reputation and strategic thinking to strengthen the company's brand and mission
Most of METR’s time horizon measurements are done using two scaffolds: Triframe and ReAct1. People sometimes see that we use these two scaffolds and feel skeptical…
The Batch AI News and Insights: I recently spoke at the Sundance Film Festival on a panel about AI.
OpenAI's GPT-5.3-Codex-Spark, an ultra-fast model optimized for real-time coding, is now available in Windsurf's Arena Mode Fast and Hybrid battle groups.
Use the Pinecone Plugin for Claude Code to develop AI Applications Faster
Meta's DINOv2 model is enhancing reforestation efforts around the world. Learn how the UK government is using DINO to help reduce costs and increase access to…
Claude Opus 4.6 (fast mode) is now available in Windsurf with limited-time promotional pricing for self serve users: 10x credits without thinking and 12x credits with…
The Batch AI News and Insights: Job seekers in the U.S. and many other nations face a tough environment.
RT CalebWe recently shipped a Data Exploration Agent in the @huggingface Dataset Viewer 💽• powered by @OpenAI gpt-oss-120b 🤖• served by @GroqInc for super fast inference…
Claude Opus 4.6 is now available in Windsurf with limited-time promotional pricing for self serve users: 2x credits without thinking and 3x credits with thinking. Available…
AnnouncementsIntroducing Claude Opus 4.6Feb 5, 2026We`re upgrading our smartest model.The new Claude Opus 4.6 improves on its predecessor`s coding skills. It plans more carefully, sustains agentic…
Infrastructure configuration can swing agentic coding benchmarks by several percentage points—sometimes more than the leaderboard gap between top models.nn
We tasked Opus 4.6 using agent teams to build a C Compiler, and then (mostly) walked away. Here's what it taught us about the future of…
China’s Moonshot releases a new open source model Kimi K2.5 and a coding agent, Google Brings Genie 3’s Interactive World-Building Prototype to AI Ultra Subscribers, and…
AnnouncementsClaude is a space to thinkFeb 4, 2026There are many good places for advertising. A conversation with Claude is not one of them.Advertising drives competition, helps…
Precision diarization, real-time transcription, and a new audio playground.
Merge pull request #558 from Keytoyze/main Update tech report
Merge pull request #557 from QwenLM/cyente-patch-4 Update README.md
Merge pull request #556 from QwenLM/add_qwen3-coder-next-info qwen3-coder-next released
Merge pull request #555 from wenting-zhao/main add qwen3 coder next tech report
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.