arXiv cs.LG
· Papers
Long-term User Engagement Optimization through Model-agnostic Downstream Rewards Learning
arXiv:2607.14192v1 Announce Type: new Abstract: As recommender systems mature in the past few years, their optimization objectives have evolved from a primary focusing on short-term behavioral signals to a broader emphasis on long-term user engagement and retention. However, directly optimizing retention is difficult b