I wonder if @xenocosmography went to watch
I wonder if @xenocosmography went to watchZhengxiao Han: Shanghai WAIC 2026
I wonder if @xenocosmography went to watchZhengxiao Han: Shanghai WAIC 2026
submitted by /u/asankhs [link] [comments]
RT Bryan Cheong大道之行也天下为公Zephyr: @teortaxesTex Bro..
Despite strong capabilities in data understanding and decision-making, autonomous data science agents still heavily rely on trial-and-error workflows that involve expensive computation. This bottleneck motivates models…
Hyper-Connections (HC) expand the residual stream of Transformers into N parallel streams, providing a form of memory scaling beyond model width and depth. Manifold-Constrained HC (mHC)…
Under model--harness co-evolution, harnesses are not merely inference-time scaffolds but data-generating components whose execution traces can shape future foundation models. This motivates harness-in-the-loop learning: optimizing harnesses…
Healthcare spans high-stakes communication, expert reasoning, and workflow execution, yet specialized LLMs that cover these use cases together remain limited. A healthcare model must handle patient…
We present Loopie, the most powerful looped Transformer to date. The Loopie series consists of two Mixture-of-Experts (MoE) models: a 20B-parameter model with 2B active parameters…
We present Audio-Visual Flamingo (AV-Flamingo), a fully open state-of-the-art audio-visual large language model (AV-LLM) for joint understanding and reasoning over audio, images, and long-form videos. Unlike…
Code review helps maintain software quality before code integration, but it also imposes a substantial workload on human reviewers. As generative artificial intelligence becomes part of…
Article URL: https://xcancel.com/__alpoge__/status/2079028340955197566 Comments URL: https://news.ycombinator.com/item?id=48973869 Points: 17 # Comments: 4
too much wordcelism around Dean's DecelgateI want estimates of 1) how much profits on inference do "The Labs" (don't fucking call them labs please, these are…
[This is an introductory blog for the paper Laguerre Geometry for Interpreting Large Language Models and the GitHub repository Geometric Lens.]LLM Lens: What does an internal…
Article URL: https://www.nytimes.com/2026/07/09/us/data-centers-native-american-tribes.html Comments URL: https://news.ycombinator.com/item?id=48973782 Points: 16 # Comments: 3
I feel like my timeline was right and now it is 3.5 months later.Assuming the Chinese government will still be okay with releasing open Mythos-class models…
RT Sakana AIRe “In Search of the Ingredients of Open-Endedness: Replicating Picbreeder with Large Vision-Language Models” won the Best Paper Award in the Complex Systems track…
Article URL: https://privacy.ca.gov/ Comments URL: https://news.ycombinator.com/item?id=48973715 Points: 5 # Comments: 0
The contest deadline was today. Here's what I submitted. Original here.Contest submission: Epistack-HowTruthfulThis is a submission to FLF's Epistemic Case Study Competition.If you're not a contest…
One of the weird things about AI is some models just turn out to be much better than others and then the next model in line…
Fable, Sol Pro, Kimi K3: "write me a short but good poem using the Odyssey as a basis, think Tennyson or Cavafy"I think this is a…
Large language models (LLMs) are transforming recommender systems from matching co-occurrence patterns in historical behavior toward reasoning about the intent that drives it. RecGPT-V1 pioneered this…
about to write a bear scenario for DeepSeek and won't even get a few millions from Dario in return. pure love of truth
Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is largely studied in isolation from the…
A community developer fine-tuned OpenBMB's MiniCPM5-1B on Claude Fable 5 traces into a 1B model that runs fully local — a 657MB smallest build, 128K context,…