Ergo: Long Form Philosophy Lectures
Article URL: https://ergo.org/ Comments URL: https://news.ycombinator.com/item?id=48840497 Points: 5 # Comments: 0
Every story across every category, newest first. Each card links to the original publisher; daily-brief posts open as editorial pages.
Article URL: https://ergo.org/ Comments URL: https://news.ycombinator.com/item?id=48840497 Points: 5 # Comments: 0
Generalist robot manipulation policies have advanced rapidly, yet existing benchmarks remain limited in systematically evaluating their capabilities. Many rely on simple, short-horizon, or skill-narrow tasks with…
RT GrokUse Grok 4.5 to build full-stack apps with ConvexMikeysee: Wow Grok 4.5 is very impressive at @convex code. Almost perfect score with a very very…
Mainstream Vision-Language-Action (VLA) models predict actions primarily from the current observation under a Markovian assumption, thus struggling with long-horizon, temporally dependent tasks. Existing memory-augmented VLAs either…
Humans can navigate an unfamiliar city and gradually form a coherent spatial mental map spanning tens of square kilometers. Can AI build spatial representations at a…
We present LingBot-World 2.0 (also known as LingBot-World-Infinity), an advanced iteration of LingBot-World featuring four distinct upgrades. (1) Our model achieves an unbounded interaction horizon while…
it surely doesnteric: 5.6 solves depression
🫶dax: i've never hyped a model release, we're generally conservative with how we use these thingsbut gpt-5.6 has had a massive impact on our team, we're…
i do love rottweilersPeter Gostev: My view of: Fable 5 vs GPT-5.6-Sol. They are not easy models to compare, these are my vibes - take them…
RT Artificial AnalysisSpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8…
Hi, I've been running vLLM on my MI50 because of tensor-parallel support. It works, but I have some complaints. For one, the quants seem much harder…
I just pushed a new audio.cpp update with streaming support and 4 ASR models: Nemotron 3.5 ASR, Higgs Audio STT, VibeVoice ASR, and Hviske ASR (da…
RT Kun Chengrok 4.5 made me give grok build a serious run todayhere's my honest first impression (non affiliated neutral view point):1. grok build is a…
Despite the recent promise in robot control, video generative models suffer from a domain mismatch due to their primary focus on content creation. For example, their…
RT grimGrok Build mogs Claude Code & Codex TUIs i can't lie. Not to mention the perf of Grok 4.5...It's not even a competition at this…
RT Akshat BubnaFun conversation with @swyx on our journey building the cloud for true elastic inference, sandboxes, and more. And of course, how we're evolving Modal's…
Article URL: https://www.alecscollon.com/blog/llm-burnout/ Comments URL: https://news.ycombinator.com/item?id=48839984 Points: 9 # Comments: 0
This paper (which appears to be going viral) is not real as far as I can tell. Faking sources has become incredibly easy thanks to good…
https://artificialanalysis.ai/evaluations/artificial-analysis-openness-index In case you want to support openness, some models are more open than others. Update: K2 think v2 is rated highest because it supplies its…
JD.com, one of the world's largest e-commerce platforms, serves over 700 million active users and millions of merchants, with a catalog of tens of billions of…
Modern one-step diffusion models achieve impressive quality through distribution-based timestep distillation. Yet, they rely on a critical assumption: Teacher and Student must inhabit the same latent…
This is the first entry in a sequence of posts which compare a mathematical theory of attention against trained transformers.Links: [GitHub repository]: The code for these…
RT Chris SweeneySometimes I still lol that @SakanaAILabs shared this example of the AI Scientist letting its intrusive thoughts win. We've all been there
RT Chris SweeneySometimes I still lol that @SakanaAILabs shared this example of the AI Scientist letting its intrusive thoughts win. We've all been there