Skip to content
arXiv stat.ML · Papers

Latent Utility Q-Learning for Preference-Adaptive Dynamic Treatment Regimes

arXiv:2307.12022v3 Announce Type: replace Abstract: Optimizing individualized treatment sequences for patients who weigh multiple, competing outcomes differently poses a challenge for dynamic treatment regime (DTR) methods, which typically assume a single univariate outcome. We propose Latent Utility Q-Learning (LUQ-Le