arXiv cs.AI
· Papers
A Contextual-Bandit Oversight Game with Two-Sided Informational Asymmetry
arXiv:2607.00155v1 Announce Type: new Abstract: We study runtime human oversight of an AI agent when private information runs in both directions: the human privately knows her reward function, while the AI privately knows the quality of the action it proposes. This is the kind of asymmetry that arises naturally when an