Funding update
In the last 6 months, METR raised commitments of around $71 million. This will fund ambitious projects: studying autonomous capabilities, tracking recursive self-improvement, evaluating monitoring systems,…
Every story across every category, newest first. Each card links to the original publisher; daily-brief posts open as editorial pages.
In the last 6 months, METR raised commitments of around $71 million. This will fund ambitious projects: studying autonomous capabilities, tracking recursive self-improvement, evaluating monitoring systems,…
.post-content .discovery-note { --discovery-font: system-ui, -apple-system, "Segoe UI", Roboto, "Helvetica Neue", "Noto Sans", "Liberation Sans", Arial, sans-serif, "Apple Color Emoji", "Segoe UI Emoji", "Segoe UI Symbol",…
Each and every open lab is working on this. Moonshot derisked Muon, DeepSeek did the most architecture R&D, ZAI seems to be at the frontier of…
sycl: fuse the gated-delta-net state writeback cpy (#26643) Port of #23940. Arc Pro B70, Qwen 3.6 27B Q4_K - Medium (48 of its 64 blocks run…
LMAO they explicitly say "we just scaled RL"That's enough. Not just Slime. DSec and Kimi's internal frameworks are all Enough. RL to superhuman software engineering in…
Open models are improving at an unprecedented rate. Congrats to the whole @Zai_org team. Ollama will support GLM-5.3 upon the open release! ❤️Get ready to code.Z.ai:…
Same prompt tested on Deepseek V4 Flash tested with codex, pi, opencode, maki, jcode, harnesses submitted by /u/Decent-Hat-5807 [link] [comments]
China is starting to harden, courtesy of Tsinghua.The US is going in on the attack. ("Criminal entities" are, for example, labs doing distillation, GPU smugglers… it's…
How to use @botDebbie O'Brien: I have a new chief of staff: In @bot I simply put: take a look at me and what i do.…
Congrats to the team!
After ≈3 hours with no feedback except "stop doing retarded tests!", DSH + V4-Pro Max had written me a pretty decent Noita 3D. Most of the…
RT ClineGLM-5.3 has released, beating Fable & the new DeepSeek V4-Pro 0813 from just yesterday on Terminal-Bench.It used GLM-5.2 as the base model with gains from…
Article URL: https://www.elttam.com/blog/ruby-4-0-universal-rce-deserialization-gadget-chain Comments URL: https://news.ycombinator.com/item?id=49295238 Points: 4 # Comments: 0
Overnight in AI: DeepSeek shipped V4-Pro and open-sourced its agent framework; OpenAI previewed 14x faster inference; and Grok 4.6 claimed the top of multiple benchmarks.
Talking-video character replacement requires coordinated transfer of appearance and voice while preserving the source motion, scene, linguistic content, and audio-video timing. Existing methods use separately optimized…
Pose-driven human animation synthesizes a video of a target person from a single reference image and a driving pose stream. Real-time generation is essential for interactive…
Article URL: https://www.mprnews.org/story/2026/08/13/attorney-says-homeland-security-spied-on-minnesotans-who-opposed-ice Comments URL: https://news.ycombinator.com/item?id=49295163 Points: 15 # Comments: 0
dflash : clarify output logging of target_layer_ids (#27013) This commit tries to make the logging of target_layer_ids a bit clearer and easier to read. Currently the…
Cactus Compute released Needle 2, an open 45M-parameter model for tool calling, device use, and structured extraction. The full model is a single 14MB binary that…
**Z.ai launched GLM-5.3**, a coding- and cyber-focused model with significant gains on agentic and security benchmarks, achieved through scaled post-training rather than a larger base model.…
Grok 4.6 ranks #1 on CursorBench for real-world codingX Freeze: Grok 4.6 just ranked #1 on CursorBench 3.2Outperforming Claude Fable 5, Opus 5 and GPT-5.6 Sol…
I didn't think we'd get here so quickly. I can run this shit on a computer I spent less than $2k for (back before prices exploded).…
Grok ImagineDéborah: Automatic animation from Grok Imagine.I didn't ask for text; it suggested that itself. The result is stunning.
oh brotherZ.ai: Introducing GLM-5.3: Built to Code. Ready for Cyber Defense.- Top-tier coding and agentic capabilities, achieved through post-training on the 743B base model- A major…