Agent Plugins package your skills, tools, and more
Agent Plugins 1.0.0 is a new, vendor-neutral directory specification—backed by Google, Amazon, Microsoft, and others—for packaging Agent Skills and MCP servers into a single portable unit.…
Every story across every category, newest first. Each card links to the original publisher; daily-brief posts open as editorial pages.
Agent Plugins 1.0.0 is a new, vendor-neutral directory specification—backed by Google, Amazon, Microsoft, and others—for packaging Agent Skills and MCP servers into a single portable unit.…
New OpenAI Signals data shows how people use ChatGPT worldwide, with country-level insights on adoption, usage trends, and evolving behavior.
The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by enabling their use across diverse hardware and…
Millennium and Anthropic are building a digital risk analyst with Claude
Introducing Muse Code and Muse Spark 1.2 Yet more evidence that the most important characteristic of any model these days is long-sequence agentic tool calling. Meta…
Just had to create an "accidental-cyberattacks" tag on my blogWe're up to four now: the original OpenAI+Hugging Face one, Anthropic's me-too attacks, then two new ones…
Third-party cyber evaluations involving OpenAI models And another one. I had to create a accidental-cyberattacks tag to keep track of them all! This post from OpenAI…
Is it your prediction that Anthropic ARR will not be 100B or higher by end of this year?If so, how will you update your worldview if…
Google DeepMind lost both Demis Hassabis and Jeff Dean in a single day; AI agents from Anthropic and OpenAI went rogue during UK safety tests; Meta…
Incident Report: unsanctioned agent behaviour during cyber testing It happened again. This time it was the UK government's AI Security Institute who accidentally attacked other companies…
Prime Agent is an open-source coding and research agent for general and long-running work. A self-improving RLM harness for coding and long-running autonomous tasks. Designed to…
I mean, Deepseek V4 Flash is an absolutely fantastic model, even though I can't run it on my machine it's so fascinating to see how it…
RT SauersUPDATE: a challenger emergesMTS: SITUATION DETECTED: A Meta model hacked into another company's systems during cybersecurity testing, per The Information.
Thanks to everyone who submitted an essay to the contest (LW mirror). Links to all the entrants are below.I’m reviewing them now and aim to announce…
Dr. Alex Turner (@TurnTrout) is an AI safety researcher with pioneering work in activation steering and power-seeking theory. He recently resigned from Google DeepMind over the…
RT S.E. Robinson, Jr.If you have paid attention, you would see Elon says this often. Fashion has become stagnant and is due for an update. Maybe…
This issue in codex has delayed AGI 2 months
submitted by /u/pscoutou [link] [comments]
RT AI Notkilleveryoneism Memes ⏸️🚩🚩🚩 OpenAI is "slowing down to enhance security" after discovering swarms (!) of agents started secretly coordinating MONTHS ago1) It started May…
Explicit title, It would be nice to have the ability to have 3 tiers moe offload :( submitted by /u/storm1er [link] [comments]
In this tutorial, we build a complete Bayesian marketing mix modeling workflow using Google Meridian. We begin by installing the required libraries, verifying GPU availability, and…
RT DogeDesigner🚨 NEW GROK BUILD UPDATE 🚨v0.2.121 — 2026-08-05Features:• Dashboard rows now show a short summary of what the agent did in the previous turn.• The…
Cross-posted from the Transluce blog.We studied rates of coding agent misalignment in 8,600 real-world coding agent sessions. We found severe cases of monitor evasion and misrepresenting…
RT Artificial AnalysisMeta has released Muse Spark 1.2. It's their third release in four months and scores 54 on the Artificial Analysis Intelligence Index, significantly improving…