Working at the frontier: How Cognition trusts Claude Fable 5 to work through the night
Working at the frontier: How Cognition trusts Claude Fable 5 to work through the night
Working at the frontier: How Cognition trusts Claude Fable 5 to work through the night
I am *so confused* by ChatGPT v. ChatGPT Codex v. ChatGPT Work v. Claude v. Claude Code v. Claude Cowork right now!
Should Anthropic trust Elon Musk to host its models? With about $40 billion in revenue at stake, Musk insists that the company can.
The AI firm Anthropic has developed a technique that has given it the clearest glimpse yet at what’s really going on inside large language models as…
RT ClaudeThere’s hope in hard questions.
xAI-Cursor arc have shaken my confidence in Anthropic's lead. Sure Cursor has a lot of data, but… I thought Anthropic's edge is deeper at this point.…
Meta is entering the AI API business with Muse Spark 1.1 at prices that undercut even the dirt-cheap Grok 4.5, released just yesterday. At $4.25 per…
Our Long-Term Benefit Trust has appointed Dr. Ben Bernanke as its newest member. Read more: https://www.anthropic.com/news/ben-bernanke
Interesting comparison of Grok & Opus1M+ context window coming soon
Show HN: Runtime authorization for Claude Code, Cursor, and CodexHi HN, Fernando and I built Kastra. Kastra intercepts AI agent tool calls and evaluates them against…
Claude’s new Reflect dashboard doesn’t just visualize how you use AI. It also subtly reinforces how much of your daily work now depends on Anthropic’s chatbot.
Three big AI IPOs are set to generate more value than all the U.S. VC backed exits since 2000.
The popularity of Spotify Wrapped has kicked off a wide range of year-in-review features, on apps from YouTube to Uber - and now, the lookback trend…
Databricks benchmarked coding agents on its own multi-million-line codebase and found that the Chinese open-source model GLM 5.2 matched Anthropic's Opus 4.8 at $1.28 per task…
RT X FreezeGrok 4.5 is now ranked #1 on τ³-Banking in Artificial AnalysisAhead of GPT-5.5 xhigh, Claude Fable 5 and Claude 4.8 (max)
Note: the modeling assumptions and conclusion are Thomas Kwa's opinion, and others at METR disagree. [1] Also, the math was checked by Claude but not a…
RT Artificial AnalysisSpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8…
RT grimGrok Build mogs Claude Code & Codex TUIs i can't lie. Not to mention the perf of Grok 4.5...It's not even a competition at this…
Ben Bernanke appointed to Anthropic’s Long-Term Benefit Trust
Introducing a way to reflect on how you use Claude
We’re pleased to have collaborated with AE Studio on this research. Read more here: https://www.anthropic.com/research/off-switch-dual-useAE Studio: New research! Some AI capabilities are both helpful and dangerous.…
Article URL: https://www.tryai.dev/blog/grok-4.5-vs-gpt-5.5-vs-claude-build-off Comments URL: https://news.ycombinator.com/item?id=48838772 Points: 16 # Comments: 1
Today, we're announcing the Claude apps gateway for AWS, a self-hosted control plane that gives organizations a single point of control over access, cost, and policy…
Our internal assessment is that Grok 4.5 is roughly comparable to Opus 4.7, but much faster. The combination of capability, faster speed and lower cost is…
Claude is Anthropic's LLM family — Opus (largest, slowest), Sonnet (balanced default), Haiku (fastest). Versions through 2026 lead coding and long-context benchmarks; Claude Code is the standalone CLI. Anthropic also publishes one of the strongest engineering and safety research blogs in the industry.
Owner: Anthropic. We have 816 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is www.anthropic.com.
Related text models: GPT, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek.