AI Daily Brief — 29 November 2025
DeepSeek releases Math-V2 — first open model with IMO gold (5/6 IMO 2025 problems, 118/120 Putnam 2024 beating top human 90). Self-verifying reasoning. AI shopping +500-700%…
DeepSeek releases Math-V2 — first open model with IMO gold (5/6 IMO 2025 problems, 118/120 Putnam 2024 beating top human 90). Self-verifying reasoning. AI shopping +500-700%…
Black Friday smashes records: $11.8B US online (+9.1%), $14.2B global AI-driven (Salesforce). AI traffic to US retail +805% YoY. AI conversions +38% Black Friday / +54%…
US Thanksgiving — frontier-lab desk dark. arXiv 'Irresponsible AI: big tech's influence on AI research and associated impacts' position paper lands. OpenAI ChatGPT-logs 20M discovery order…
Light pre-Thanksgiving cycle: Perplexity + PayPal Instant Buy extends with Black Friday merchant list. TSMC: advanced-node capacity ~3× short of demand (2nm booked through 2028, NVIDIA…
Judge Alsup grills ClaimsHero CEO Freund at 4-hour Bartz hearing — warns of criminal referral over 'bait-and-switch' opt-out pitch. Anthropic publishes Opus 4.5 prompt-injection research (~1%…
Anthropic ships Claude Opus 4.5 — first model to break 80% SWE-bench Verified (80.9% vs Gemini 3 Pro 76.2% / GPT-5.1 76.3%); OSWorld 66.3%. $5/$25 per…
Sunday before the storm: Claude Opus 4.5 confirmed for Mon Nov 24 — claims SOTA SWE-bench Verified ahead of Gemini 3 Pro + GPT-5.1; $5/$25 per…
Quiet Saturday: Willison deep-dive on Ai2 Olmo 3 (full transparency). HunyuanVideo-1.5 dominates HF + GitHub weekend traffic. Anthropic publicly quiet before Mon Opus 4.5 reveal (model…
SC25 closes — El Capitan #1 (1.809 EF/s HPL + 16.7 EF/s HPL-MxP), JUPITER first non-US exascale. NVIDIA Apollo open AI-physics family + RIKEN GB200 supercomputers.…
Google launches Nano Banana Pro (Gemini 3 Pro Image) — 94% text-rendering, 2K/4K, SynthID, across Gemini + Workspace + Ads + AI Studio + Vertex. Genspark…
EU 'Digital Omnibus on AI' published — defers high-risk AI obligations to Dec 2027 / Aug 2028. Draft Trump EO challenging state AI laws leaks. Ignite…
Google ships Gemini 3 Pro + Deep Think (HLE 37.5/41%, GPQA 91.9/93.8%, AIME 95%, LMArena 1501; ARC-AGI-2 45.1%) + Antigravity agentic IDE + generative UI. Microsoft…
xAI ships Grok 4.1 silently — #2/#3 LMArena, #1 EQ-Bench3, hallucination 12.09% → 4.22% (free for all). TOP500: El Capitan #1 (1.809 EF/s), JUPITER becomes first…
SC25 opens in St. Louis (16,500+ attendees, record 524 exhibitors) — Sunday tutorials + workshops + exhibitor reception. NVIDIA + AMD + Intel + ARM stage…
Quiet Saturday: $3.7B AI funding week (Anysphere $2.3B + CHAOS Industries $510M + d-Matrix $275M + Gopuff $250M + Alembic $145M = 96% to 5 winners).…
GPT-5.1 + Cursor mega-round + Baidu ERNIE 5.0 set agenda heading into SC25 + Microsoft Ignite week. Munich Court rules OpenAI training on German lyrics infringes…
Cursor closes $2.3B Series D at $29.3B (Accel + Coatue, with NVIDIA + Google + a16z + Thrive); $1B+ ARR, 300+ employees, Composer agentic coding model…
OpenAI launches GPT-5.1 (Instant + Thinking) with adaptive reasoning + 8 personality presets; keeps GPT-5 available 3 months. Anthropic confirms $50B US infrastructure with Fluidstack (TX…
Yang Zhilin Reddit AMA — walks back $4.6M K2 Thinking cost as 'not official', confirms H800 + Infiniband training, previews K3 with KDA architecture + possible…
Quieter Monday: Scribe hits $1.3B unicorn valuation with $75M Series C (78K paid orgs, 94% of Fortune 500). AirOps $40M Series B at $225M for AI…
Quiet Sunday: Kimi K2 Thinking aftermath continues with 'closest open models to closed frontier' framing. Willison's tau-squared-Bench Telecom 93% score (highest agentic tool-use he's measured). OpenAI-AWS…
Quiet Saturday — Kimi K2 Thinking aftermath dominates with 'DeepSeek moment' analyses, $4.6M training-cost claim ricocheting, Lambert / LessWrong / Barnacle Goose breakdowns. $2.3B weekly AI…
Microsoft Copilot Fall Release — 'Mico' assistant + Real Talk style + Memory + Groups + Imagine + AI Search with clickable citations. Seven new wrongful-death/product-liability…
Moonshot AI releases Kimi K2 Thinking — 1T MoE / 32B active / INT4 native / 256K / 200-300 sequential tool calls / ~$4.6M training /…