arXiv stat.ML
· Papers
Latent Utility Q-Learning for Preference-Adaptive Dynamic Treatment Regimes
arXiv:2307.12022v3 Announce Type: replace Abstract: Optimizing individualized treatment sequences for patients who weigh multiple, competing outcomes differently poses a challenge for dynamic treatment regime (DTR) methods, which typically assume a single univariate outcome. We propose Latent Utility Q-Learning (LUQ-Le