LessWrong AI
· Communities
Synthetic Scalable Oversight
We propose synthetic scalable oversight, a technique for studying scalable oversight by creating graphical abstractions of real-world problems and training tiny models inside these synthetic environments as a proxy for training LLMs at scale.We thank Oliver Richardson, Christian Szegedy, Michael Douglas, and countless