Skip to content
arXiv cs.LG · Papers

Long-term User Engagement Optimization through Model-agnostic Downstream Rewards Learning

arXiv:2607.14192v1 Announce Type: new Abstract: As recommender systems mature in the past few years, their optimization objectives have evolved from a primary focusing on short-term behavioral signals to a broader emphasis on long-term user engagement and retention. However, directly optimizing retention is difficult b