Skip to content
X · @OpenAI · X / Twitter

RT Noam Brown: Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations …

RT Noam BrownLong-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss.We’re sharing what we learned from studying a long-running model, and how those findings are shaping our approach to evaluations, alignment, monitoring, and user contr