LWiAI Podcast #232 – ChatGPT Ads, Thinking Machines Drama, STEM
OpenAI to test ads in ChatGPT as it burns through billions, The Drama at Thinking Machines, STEM: Scaling Transformers with Embedding Modules
Every story across every category, newest first. Each card links to the original publisher; daily-brief posts open as editorial pages.
OpenAI to test ads in ChatGPT as it burns through billions, The Drama at Thinking Machines, STEM: Scaling Transformers with Embedding Modules
Building the “young person’s AGI lab” to unlock data efficient models, which we believe is the bottleneck to laddering up the next rung of AI intelligence
Pinecone Assistant Node in n8n: Turn Any Data Source Into Knowledge
State-of-the-art video generation across quality, cost, and latency.
The Information reports Anthropic raises 2026 revenue forecast 20% to $18B (2027: $55B). Moonshot ships Kimi K2.5 — open 1T-parameter MoE with Agent Swarm of up…
Thriving in a world of agents
Why vector databases are here to stay.
Massive Monday: NVIDIA invests $2B in CoreWeave at $87.20/share for 5+ GW of AI factories by 2030. Microsoft unveils Maia 200 inference chip (TSMC 3nm, 216…
Quiet Sunday. Crescendo AI's weekly aggregator reports an early OpenAI GPT-5.3-Codex disclosure — to be confirmed by OpenAI release notes.
Quiet Saturday — week shaped by Baidu ERNIE 5.0 (Jan 22) and Qwen 3 Max Thinking (Jan 23). Industry attention turns to next-week's Maia 200 /…
And an Overview of Recent Inference-Scaling Papers
Alibaba kicks off an unusually fast Qwen cadence with Qwen 3 Max Thinking (snapshot qwen3-max-2026-01-23). Mastercard's Agent Pay integrates with Microsoft Copilot Checkout and OpenAI Instant…
OpenAI to test ads in ChatGPT as it burns through billions, Sequoia to invest in Anthropic, Zhipu AI breaks US chip reliance, The Drama at Thinking…
The Batch AI News and Insights: How can businesses go beyond using AI for incremental efficiency gains to create transformative impact?
ollama launch is a new command which sets up and runs coding tools like Claude Code, OpenCode, and Codex with local or cloud models. No environment…
Baidu ships ERNIE 5.0 — a 2.4-trillion-parameter ultra-sparse MoE that processes text, image, audio and video. Ranks #1 among Chinese models and #8 globally on LMArena,…
Railway, a San Francisco-based cloud platform that has quietly amassed two million developers without spending a dollar on marketing, announced Thursday that it raised $100 million…
Kais, Blockit, and Optimizing Time
Anthropic Engineering publishes 'Designing AI-resistant technical evaluations' — Tristan Hume walks through three iterations of the company's performance-engineering take-home test, each defeated by successive Claude versions.
Heaps do lie: debugging a memory leak in vLLM.
What we learned from three iterations of a performance engineering take-home that Claude keeps beating.
Hugging Face publishes 'One Year Since the DeepSeek Moment' — DeepSeek R1 most-liked model in HF history; Chinese labs now hold ~15% of global model share…
Merge pull request #247 from EAGzzyCSL/zzy/link-midscene add quick start guide for Midscene
Today, we are releasing LFM2.5-1.2B-Thinking, a reasoning model that runs entirely on-device. It fits within 900 MB of memory on a phone and delivers both the…