Skip to content
arXiv cs.AI · Papers

ProDVI: Programmatic Dynamics Priors for Value Network Initialization

arXiv:2608.06015v1 Announce Type: cross Abstract: Deep Reinforcement Learning (RL) is notoriously sample inefficient. One contributing factor is that RL agents are typically initialized from scratch, forcing them to acquire task-relevant knowledge through online interaction. Existing approaches obtain informative initi