LessWrong AI
· Communities
Sub-agent delegation chaining
Epistemic status: pretty confident in the validity of the core proposal, not that confident in specific implementation detailsTL;DR: we should cryptographically verify that sub-agent instances/sessions are downstream of human instructionsFrontier AI labs have started using LLM-based monitoring systems to check for misb