The mistake of conflating intelligence and power
f this is your definition of intelligence is the ability to achieve your goals across a wide variety of domains, then Stalin was the most intelligent…
f this is your definition of intelligence is the ability to achieve your goals across a wide variety of domains, then Stalin was the most intelligent…
the verification loop for theories can be on the order of decades and centuries, and even then we know today as the better theory can often…
An eventful month with one flagship release after another
New article: a visual tour of recent LLM architecture advances, from Gemma 4 to DeepSeek V4.I focus on long-context efficiency tweaks like KV sharing, per-layer embeddings,…
From Gemma 4 to DeepSeek V4, How New Open-Weight LLMs Are Reducing Long-Context Costs
Google DeepMind and Singapore partner to apply frontier AI to address complex challenges across health, education, and sustainability and more.
Clare Bryant uses Co-Scientist to identify genetic triggers in emerging infectious diseases.
Calico Life Sciences uses Co-Scientist to connect scattered findings and generate new leads in aging research.
Filippo Menolascina uses Co-Scientist to identify new liver disease treatments and explain why existing drugs only help certain patients.
Co-Scientist unites Boston Children’s Hospital and MIT’s labs to explore new RNA-based treatments for ALS.
Stanford geneticist uses Co-Scientist to help find new treatments for chronic liver disease and liver fibrosis.
Learn how our WeatherNext AI model help forecasters give communities unprecedented time to prepare ahead of the historic Hurricane Melissa.
Gemini 3.5 launches: frontier intelligence with action; Microsoft Research on delegation/long-horizon reliability; The Batch on China + Meta.
Gemini 3.5 is built to help you execute complex, agentic workflows.
Our recent paper, “LLMs Corrupt Your Documents When You Delegate”, has generated discussion about the reliability of AI systems in delegated workflows. We appreciate the interest…
Hottest AI job: FDEs / Trump in China / No jobpocalypse / AI models fixing benchmarks / Americans agree: "no datacenters here" / Claude 3 vs…
AlphaGo is still the cleanest worked example of the primitives of intelligence: search, learning from experience, and self-play.
A new scaling law that relates particular architectural choices to loss helps identify models that improve throughput by up to 47% with no loss of accuracy.
**Cerebras** made headlines with its **IPO**, marking a significant milestone for the company known for its contrarian hardware approach. The **Cerebras CFO Bob Komin** emphasized the…
Together AI partners with Pearl Research Labs to launch a discounted Pearl-powered inference endpoint for Gemma-4-31B-it-pearl, using Proof of Useful Work to turn AI workloads into…
The Batch AI News and Insights: We’ve been working on AI Andrew, an AI companion shaped by my personality.
Use your Grok account and subscription inside Nous Research’s open-source, self-improving Hermes agent.
Anthropic $200M partnership with Gates Foundation; PwC deploys Claude across enterprise; Amazon Promptimus; Ng on Transformers in Practice.
New course: Transformers in Practice. You'll get a practical view of how transformer-based LLMs work, so you can reason about their behavior, diagnose problems like slow…