Skip to content
arXiv cs.LG · Papers

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary

arXiv:2607.11953v1 Announce Type: new Abstract: Does a reinforcement-learning agent that earns high reward represent its task's latent state, or only a reward-correlated shortcut? The question is usually unanswerable: the "true state" is undefined. We make it exactly answerable with a white-box instrument: express the