The Case for AI Behavioral Science
This is NOT an anthropomorphizing case.It is my view that AI research is, at the moment, split between three camps:Model-internal research (mechinterp folks)Capability research (METR, Irregular,…
This is NOT an anthropomorphizing case.It is my view that AI research is, at the moment, split between three camps:Model-internal research (mechinterp folks)Capability research (METR, Irregular,…
Article URL: https://www.dailymail.com/news/article-15943905/Mystery-identity-Green-Boots-climber-macabre-landmark-frozen-ice-dying-Everest-finally-solved-DNA-test.html Comments URL: https://news.ycombinator.com/item?id=48768336 Points: 10 # Comments: 0
Lydia Laurenson recently posted an article called "The Inside Story of Leverage Research" that gets into substantially more detail on what went on in that organization…
I'm building a website and want to automate testing of my UI. I know Claude Code and Codex can do this, but due to cost and…
RT Mada SegheteWhat an honor to curate the first AI in GTM track at @aiDotEngineer 😆 Heard that we need a bigger room next year @swyx…
Re Program Manager (RSI Lab)https://sakana.ai/careers/program-manager-rsi-lab/
I liked the idea of a nano llm, but decided to actually challenge myself with developing one. Keep in mind I developed this model, do your…
RT adlin is bts @ ai engineer wfThe star @mochipomsky is diligently listening to her favourite person @swyx
Credit to /u/Superb-Translator236 for original posting on another sub - this sub doesn’t allow cross-posting. submitted by /u/Thrumpwart [link] [comments]
We need to convince EAsians that nootropics work at boosting scores on their shitty life-breaking evals^W exams like gaokao. Surely they'll soon hill-climb to something better…
Arxiv paper: https://arxiv.org/abs/2606.03811 submitted by /u/Thrumpwart [link] [comments]
Model Y long wheelbase now available to order in the USTesla: Introducing Model Y Long Wheelbase – now available in the US & Puerto RicoA 3-row,…
Three of the most popular methods for training language models to reason look like three different tricks. They are not. All three adjust a single number:…
People overthink; language models over-sample, and the extra effort can talk both into a worse answer. Reasoning systems answer a hard question by sampling it many…
https://scalingintelligence.stanford.edu/blogs/hipkernels/ submitted by /u/Superb-Translator236 [link] [comments]
TL;DRWhen a model role-plays a persona, does it only change what it says, or also what it internally represents as true?To study this, we induce personas…
Article URL: https://codeberg.org/jjba23/modusregel Comments URL: https://news.ycombinator.com/item?id=48767834 Points: 8 # Comments: 0
RT Alex CheemaFull house for Local AI Track @aiDotEngineerWe’re going to make Local AI The DefaultStacked lineup:@nvidia x @exolabs x @OsmanticAI x @roboflow x @PrimeIntellect x…
RT Marc Andreessen 🇺🇸Turbo America. 🇺🇸🚀Rapid Response 47: .@POTUS: "There’s no reason we should stop at 4%. We should be at 12% and 13% GDP... That’s…
In this tutorial, we build a RAG-Anything workflow to explore how multimodal retrieval works across text, tables, equations, and images. We prepare a Colab environment, enter…
https://github.com/ggml-org/llama.cpp/pull/25222 Another win for Intel ARC users (all 4 of us). The community keeps improving llama.cpp for Intel ARC. This time, the hero from that Pull…
AI has transformed how organizations operate, driving unprecedented levels of productivity and innovation. However, AI adoption can be impeded by concerns...
Adobe is experimenting with “agentic sites” that generate pages around an individual user’s intent. At AIEWF, we talked to Carlos Sanchez about the Web's future.
I remember @ConradBastable teaching me to respect financialism. It is indeed one of the absurd, broken magicks of our civilization – like math, like faith. But…