Skip to content
arXiv cs.AI · Papers

A Contextual-Bandit Oversight Game with Two-Sided Informational Asymmetry

arXiv:2607.00155v1 Announce Type: new Abstract: We study runtime human oversight of an AI agent when private information runs in both directions: the human privately knows her reward function, while the AI privately knows the quality of the action it proposes. This is the kind of asymmetry that arises naturally when an