How to pace the US frontier
IntroductionLast week, the Pacing the Frontier open letter, signed by over 1,000 frontier AI employees, requested “the U.S. government support an international effort to develop the…
Every story across every category, newest first. Each card links to the original publisher; daily-brief posts open as editorial pages.
IntroductionLast week, the Pacing the Frontier open letter, signed by over 1,000 frontier AI employees, requested “the U.S. government support an international effort to develop the…
Locking in AI safety regulation now is a mistake.
RT DailyPapersNVIDIA just released the NeMo Gym conversational tool-use assets on Hugging FaceA bundle of golden policy/tool reference pairs and prompt histories for Gym's conversational tool-use…
RT Steve RattnerThis morning's jobs report was shockingly negative.Not only did the economy lose 23,000 jobs in July — the previous estimates for May and June…
I’m curious whether there is now a theoretical or empirical “sweet spot” for LLM quantization, preferably research done using open-source formats like GGUF Suppose you have…
A fresh llama.cpp PR (#26689) changes what looks like a tiny SYCL FlashAttention dispatch decision. With a quantized KV cache ("q4_0" / "q8_0"), decode was being…
RT Sebastian MallabyMy column on @demishassabis in The Times. —Given his record of caring about AI safety and impact, people should err on the side of…
Article URL: https://www.cbc.ca/news/business/canada-jobs-july-2026-9.7299225 Comments URL: https://news.ycombinator.com/item?id=49213367 Points: 15 # Comments: 1
How does the situation keep turning out to be worse than we know? How much should we update, therefore, that it is a lot worse than…
server: (router) add LRU scheduler (#26572) add lru_sched handle coalescing (req leaves waiting queue) add tests fix stream case address review comments Website: https://llama.app macOS/iOS: macOS…
Responding to the next frontier of critical cyber capabilities
TLDR:Many important decisions for safety depend on or are influenced by benchmark scores. These benchmarks, in effect, are trying to measure latent properties of models from…
To all those in the US: Are you planning to go Sydney or Atlanta this year for NeurIPS? submitted by /u/rsesrsfh [link] [comments]
Some of the biggest names on Google's AI team got new jobs this week. In some cases, including for legendary Googler Jeff Dean, those jobs are…
Article URL: https://www.lse.ac.uk/research/research-for-the-world/economics/tax-cuts-for-the-wealthy-only-benefit-the-rich-debunking-trickle-down-economics Comments URL: https://news.ycombinator.com/item?id=49213097 Points: 7 # Comments: 0
OpenAI is planning a donut-shaped smart speaker for 2027, priced above $300. The screenless device has a camera, microphones, and moving parts. It's designed to learn…
Article URL: https://sfstandard.com/pacific-standard-time/2026/08/07/sf-ai-billboards-dystopian-not-funny/ Comments URL: https://news.ycombinator.com/item?id=49212928 Points: 16 # Comments: 25
congrats to oklo for achieving criticality!(less than a year after groundbreaking)https://oklo.com/newsroom/news-details/2026/Oklos-Groves-Reactor-Achieves-First-Criticality-in-Under-a-Year/default.aspx
In this post, you learn how Cohere Health built a multi-tenant agentic architecture on AgentCore using AgentCore Runtime’s secure MicroVM isolation, unified tool access through AgentCore…
I have not been successful with management to get funding for local resources despite bringing forth solid arguments about data sovereignty and related architectures. What actually…
server: (router) do not evict busy models (#26567) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS…
TReNDS, a research center at Georgia State University, built an agentic AI pipeline on Amazon Bedrock and the open-source Strands Agents SDK that automatically investigates production…
The AWS Generative AI Innovation Center built an automated system that uses constraint programming and custom tree search to determine, with mathematical certainty, when and how…
Or is there any reason why I feel like model output quality seems to be better when I use higher micro-batch values (ub) in llama-cpp? I…