Skip to content
LessWrong AI · Communities

Synthetic Scalable Oversight

We propose synthetic scalable oversight, a technique for studying scalable oversight by creating graphical abstractions of real-world problems and training tiny models inside these synthetic environments as a proxy for training LLMs at scale.We thank Oliver Richardson, Christian Szegedy, Michael Douglas, and countless