AI safety prizes
Rather than paying up front for AI safety research (push funding), perhaps we should pay after the fact for the work that made the most progress…
Rather than paying up front for AI safety research (push funding), perhaps we should pay after the fact for the work that made the most progress…
To understand why powerful AIs might get into conflict, and ways to mitigate it, we need to understand bargaining problems: situations where multiple agents have different…
Today, I'll be linking both a Epoch blog post and a paper called The Bounded Parallelizability of R&D: Theory and Application to AI, about how standard…
Based on the academic paper: Measuring Intelligence Beyond Human Scale[1]TLDR: Historically, intelligence benchmarks have been composed of human-generated questions. However, current techniques do not scale to…
And why our 2023 regrant to Timaeus was goatedWe think the for-profit funding ecosystem has some cool properties. Different funders naturally come in at different stages,…
GDM’s AGI Safety and Alignment Team is hiring for multiple roles. This is the team at GDM, led by Rohin Shah, that aims to reduce existential…
This work was done by an automated research scaffold developed at Redwood Research. abhayesian provided the initial project idea. The agent designed and ran all experiments…
We must take great care not to ignore the things that are not easily quantified - Brian Christian, The Alignment ProblemIntroductionModel evaluations have a problem. This…
This post is meant as a background, or "relevant context", for our sequence on AI oversight and its limitations. It can also be read on its…
Thank you to Jimmy Sastra and Kitchener D. Wilson for discussions relating to this piece. All opinions are my own.What are organoids anyway?If someone placed a…