LessWrong AI
· Communities
Making benchmarks outputs directly useful for AI safety and security
Epistemic status: written in 30 min. This is not as polished as I’d like but I prefer to share this as is than not to share it at all.AI are becoming increasingly good at solving problems. Benchmarks are saturating fast.I think we should take advantage of this to make them solve useful problems while being evaluated. E