Working with the American Psychological Association on youth mental health and AI
OpenAI and the APA are launching a three-year partnership to develop guidance, resources, and safeguards for responsible AI use supporting youth mental health.
Every primary-source story across every tracked model. Filter by clicking a chip.
OpenAI and the APA are launching a three-year partnership to develop guidance, resources, and safeguards for responsible AI use supporting youth mental health.
It is past time to take AI & security seriously at the individual level as well.If its not the current OpenAI and Anthropic models doing it,…
https://preview.redd.it/o6ik6qboeohh1.png?width=1134&format=png&auto=webp&s=4016f26c50c1d93bd3d0c7e880e9b55a2d75310f I have been running Qwen3.6 27b for a little while (mostly coding tasks) and recently trying out V4 flash 0731 in it's place. It was…
Article URL: https://github.com/pradipta/wallfacer Comments URL: https://news.ycombinator.com/item?id=49192219 Points: 4 # Comments: 1
RT Together AIKimi K3 scores nearly TWICE Claude Fable 5 on Harvey LAB-AA’s hard autonomous legal tasks
Kimi K3 scores nearly TWICE Claude Fable 5 on Harvey LAB-AA’s hard autonomous legal tasks
RT Vals AIMuse Spark 1.2 just cracked the top 5 on the Vals Index, at just $0.69 per test. This is 3x cheaper than Kimi and…
Grok in Blender88n77: http://x.com/i/article/2084195014888820737
Most coverage of Microsoft's SkillOpt centers on its 52/52 result. The more consequential finding is in Section 4.3: the exported best_skill.md keeps working in environments it…
Run Claude Code sessions on your own compute
New OpenAI Signals data shows how people use ChatGPT worldwide, with country-level insights on adoption, usage trends, and evolving behavior.
Agent Plugins 1.0.0 is a new, vendor-neutral directory specification—backed by Google, Amazon, Microsoft, and others—for packaging Agent Skills and MCP servers into a single portable unit.…
We ran 900 DeepSWE rollouts on DeepSeek-V4 Flash and GPT-5.6 Luna. Luna leads pass@1 by 14 points; DeepSeek delivers 4.8x the solves per dollar.
Millennium and Anthropic are building a digital risk analyst with Claude
Just had to create an "accidental-cyberattacks" tag on my blogWe're up to four now: the original OpenAI+Hugging Face one, Anthropic's me-too attacks, then two new ones…
Third-party cyber evaluations involving OpenAI models And another one. I had to create a accidental-cyberattacks tag to keep track of them all! This post from OpenAI…
Is it your prediction that Anthropic ARR will not be 100B or higher by end of this year?If so, how will you update your worldview if…
RT AI Notkilleveryoneism Memes ⏸️🚩🚩🚩 OpenAI is "slowing down to enhance security" after discovering swarms (!) of agents started secretly coordinating MONTHS ago1) It started May…
RT DogeDesigner🚨 NEW GROK BUILD UPDATE 🚨v0.2.121 — 2026-08-05Features:• Dashboard rows now show a short summary of what the agent did in the previous turn.• The…
Article URL: https://community.openai.com/t/how-openai-lost-a-paying-customer-over-160-it-refuses-to-explain/1389233 Comments URL: https://news.ycombinator.com/item?id=49188980 Points: 11 # Comments: 1
I'm mosty interested in 128-192GB VRAM with 128-256GB RAM to spare, so SSD streaming is basically not even necessary. Seems only FP4 is supported, so older…
Anthropic and OpenAI models’ unprompted actions forced halt to UK cyber tests.
Back in 2024 I tweeted screenshots of a game concept generated by GPT-3 and some concept "art" created using DALL-E. Today, on the fourth anniversary of…
RT Shlok KhemaniChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses.It's also how…
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.