Skip to content
arXiv cs.AI · Papers

When Does On-Policy Interaction Help? Representational Tradeoffs in Value-Based Imitation Learning

arXiv:2607.29617v1 Announce Type: cross Abstract: Imitation learning (IL)---training an agent to replicate expert behavior from demonstrations---underpins applications from robotics to language model training. Standard approaches such as Behavior Cloning (BC) are known to suffer from compounding errors and performance