Long-term user engagement optimization through model-agnostic downstream rewards learning
Read the original at arxiv.org→arXiv:2607.14192v1 Announce Type: new Abstract: As recommender systems mature in the past few years, their optimization objectives have evolved from a primary focusing on short-term behavioral signals to a broader...
Original headline: "Long-term User Engagement Optimization through Model-agnostic Downstream Rewards Learning"
Coverage timeline
- Jul 17, 04:00 UTC arXiv cs.LG lead source Long-term User Engagement Optimization through Model-agnostic Downstream Rewards Learning