arXiv stat.ML
· Papers
Generalized Gaussian Temporal Difference Error for Uncertainty-aware Reinforcement Learning
arXiv:2408.02295v4 Announce Type: replace-cross Abstract: Conventional uncertainty-aware temporal difference (TD) learning often models TD errors as zero-mean Gaussian. This assumption can miss the heavy-tailed and heteroscedastic residuals induced by bootstrapping and exploration. We introduce a state-conditioned shap