And yet they still stink at good long-form fiction.
And yet they still stink at good long-form fiction.
And yet they still stink at good long-form fiction.
RT StarlinkStarlink enables fast, reliable connectivity where traditional infrastructure does not exist 🛰️🏔️OffRoadCampingBoys: @RummingBum Many of the areas we went, there is not a single cell…
RT FuserLuma Ray 3.2 from @LumaLabsAI is now live in Fuser.Keep the action intact. Reframe your scene in any aspect ratio.Try it now↓
I've been saying, for a long time. The interesting political question is whether you can arrange the promotion of virtuous people to positions of power. Everything…
Wei Lui is «cooking coding agents @deepseek_ai» btwI think we'll see 0731 added to this eval soonI also think it's getting added to their internal hill-climb…
I continue to think that a lack of verifiable answers in many fields is a real issue for LLMs but not as big a problem as…
good way to organize main chats vs /side chats:doing the work vsdoing metaworkagrim singh: @swyx i use the /side to ask the 'are you stuck' questions…
> Google was too nervous to release it and DeepMind was blocked from shipping products that could disrupt Google.bookmark for the next vc that asks you…
RT Kevin KwokSo uplifting and painful to see exactly the level of interviews we could be having but arentAnd Nolan engages so much more in response…
RT martin_casadoIt appears model capabilities / releases are accelerating. But if you divide by the money going into the labs, it looks sublinear. Clearly orgs get…
Josh had to stop @vipulved and ask him to repeat the number @wolfejosh Together AI went from serving 30B tokens a month to 400T 📈Over 10,000x…
A special effect like this used to take months of effort by a specialized company
ExactlyBrivael Le Pogam: Après toutes ces vagues d’immigration en Espagne, je vais vous expliquer par A + B pourquoi une partie de la gauche a structurellement…
Why is evaluating agents so difficult relative to evaluating a standard LLM?An LLM generates a single response to a prompt. An agent instead interacts with an…
RT Maye MuskA friend of mine said that we should welcome illegal immigrants, so long as they don’t live near her. She’s now a former friend.…
RT ZainPareto curve? More like the pareto step functionArena.ai: Exciting news: DeepSeek-V4-Flash-High by @deepseek_ai has reshaped the Pareto Frontier in the Frontend Code Arena, with a…
Grok 4.5 is Pareto #1 when considering speed & costGavin Baker: Also happy to see that that Grok 4.5 was the best frontier model in their…
Our Twitch streams with @Cohere_Labs are #1 in the tech category⚡️Cohere Labs' free ML Summer School series has brought together viewers from 20+ countries to learn…
RT NASA Administrator Jared IsaacmanNASA recently recognized one of America’s outstanding engineers. At @SpaceX’s Rocket Development Facility in McGregor, Texas, I joined Congressman @PeteSessions and McGregor…
One other observation: for almost every human on the planet, this is not just beyond our abilities but beyond our ken. We can only trust expert…
RT Yun-Ta TsaiAI will increasingly do more and more difficult tasks that would take longer and longer to comprehend.We will, at first, gasp.Then, over time, it…
LLMs are a wonderful driver for the advancement of humanity. We just need to proactively manage the transition to powerful AI and the benefits will be…
team humanityjason: one of the most beautiful things about OpenAI is that every employee really has a voice.i wanted to capture what it feels like to…
This type of work is incredibly valuable for explaining to the broader public how you can keep releasing open models despite them having open ended risks.…