connect your airtable and chatgpt:
connect your airtable and chatgpt:Airtable: We've partnered with @OpenAI to bring Airtable directly into @ChatGPT. Build fully customized workflows and manage your data without ever leaving…
Every primary-source story across every tracked model. Filter by clicking a chip.
connect your airtable and chatgpt:Airtable: We've partnered with @OpenAI to bring Airtable directly into @ChatGPT. Build fully customized workflows and manage your data without ever leaving…
And Grok 4.6 is a significant improvementben hylak: guys i really hate to break this to you but grok 4.5 high fast is actually good. i…
GPT-5.6 found optimizations that "reduced end-to-end serving costs by 20%" for OpenAI to serve that modelPresumably that's billions of dollars a month in savings at this…
I feel more confidence in DeepSeek's "we will delay V4 until we co-optimize the model and the harness enough" decision now. Cloning Claude Code won't cut…
Rather, I interpret Slopnet 5 and Slopus 5 as more evidence for my “post-consumer market” hypothesis. Anthropic openly despises B2C. All resources going towards Mythos for…
RT zhyncsInteresting data from the OpenRouter Kimi K3 dashboard:@togethercompute offers one of the lowest prices for @Kimi_Moonshot K3 while also achieving one of the highest prompt…
Microsoft pitched its own homegrown AI models, harnesses, and even a Mythos competitor on Wednesday, telling Wall Street it plans for continued growth.
I have been pushing up to 90 commits a day on a MacBook Air via 4-5 parallel agents. As you can imagine when all the agents…
Google's open-source TPU microbenchmark suite provides developers with granular performance metrics across Network, Compute, HBM, Host Transfer, and Attention components to validate real-world hardware capabilities. By…
avatarin uses OpenAI’s GPT-Realtime to give Yamada Denki shoppers 24/7 multilingual support. In two weeks, 30,000 people used the agent and 92% of survey responses were…
Investigating three real-world incidents in our cybersecurity evaluations
This research was done as my capstone project during ARBOx4.Epistemic Status: I'm relatively sure the results I obtained and my interpretations are correct. I'm unsure if…
Nobody is ahead of OpenAI on post-trainingTibo: Turns out GPT-5.6 Sol is actually SoTA on ARC-AGI-3. Just took two setting changes. You just have to allow…
Has anyone yet tried to extract experts from kimi (or GLM 5.2) per chance? There is REAP that removes experts based on routing, but I could…
How enabling two settings tripled our scores on the ARC-AGI-3 benchmark
When Microsoft reported killer fourth-quarter earnings for its fiscal 2026 year (which ended June 30), it tucked in an interesting little tidbit about how its investments…
we're experimenting with our own dynamic GGUF quants of kimi k3, made from the original weights with our llama.cpp fork. Q3_K_S is done and works 1114.76…
Periodic reminder that OpenAI has insane marginsOpenAI: After deployment, we applied GPT-5.6 Sol to advance the frontier of efficiency by making itself more efficient to run.…
ChatGPT Voice for spending more time working from where you want:Alex Finn: ChatGPT Voice has 100% transformed how I workInstead of spending 12+ hours a day…
GPT-5.6 Sol for improving production serving efficiency. One of the ways we're able to get such great price-performance:OpenAI: After deployment, we applied GPT-5.6 Sol to advance…
Weng previously served as the VP of AI Safety Research at OpenAI.
xAI is suing Minnesota Attorney General Keith Ellison over a law passed back in May that broadly targets "nudification" apps, claiming that the statute's punitive provisions…
I think the claims about having Mythos level-model in our laptops in 1-2 years might not be so crazy of a theory submitted by /u/SilverRegion9394 [link]…
Grok Voice is now #1 in agentic performanceArtificial Analysis: SpaceXAI has released Grok Voice Think Fast 2.0 today, with the High reasoning variant debuting at #2…
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.