Skip to content
LessWrong AI · Communities

Research update: RL on Debate Games shows Proposal Accuracy uplift alongside Judge Hacking

The first three sections are written for a general TAIS reader who wants to understand what the state of Debate research is and some high-level takeaways of our work. A reader familiar with Debate may like to skip the setup and start with our presentation of An illustrative training run. The remaining sections are writ