research-article

Open access

PrefRec: Recommender Systems with Human Preferences for Reinforcing Long-term User Engagement

Authors:

Kun Gai,

Bo AnAuthors Info & Claims

KDD '23: Proceedings of the 29th ACM SIGKDD Conference on Knowledge Discovery and Data Mining

Pages 2874 - 2884

https://doi.org/10.1145/3580305.3599473

Published: 04 August 2023 Publication History

PDF eReader

Abstract

Current advances in recommender systems have been remarkably successful in optimizing immediate engagement. However, long-term user engagement, a more desirable performance metric, remains difficult to improve. Meanwhile, recent reinforcement learning (RL) algorithms have shown their effectiveness in a variety of long-term goal optimization tasks. For this reason, RL is widely considered as a promising framework for optimizing long-term user engagement in recommendation. Though promising, the application of RL heavily relies on well-designed rewards, but designing rewards related to long-term user engagement is quite difficult. To mitigate the problem, we propose a novel paradigm, recommender systems with human preferences (or Preference-based Recommender systems), which allows RL recommender systems to learn from preferences about users' historical behaviors rather than explicitly defined rewards. Such preferences are easily accessible through techniques such as crowdsourcing, as they do not require any expert knowledge. With PrefRec, we can fully exploit the advantages of RL in optimizing long-term goals, while avoiding complex reward engineering. PrefRec uses the preferences to automatically train a reward function in an end-to-end manner. The reward function is then used to generate learning signals to train the recommendation policy. Furthermore, we design an effective optimization method for PrefRec, which uses an additional value function, expectile regression and reward model pre-training to improve the performance. We conduct experiments on a variety of long-term user engagement optimization tasks. The results show that PrefRec significantly outperforms previous state-of-the-art methods in all the tasks.

Supplementary Material

MP4 File (1148-2min-promo.mp4)

Presentation video for PrefRec: Recommender Systems with Human Preferences for Reinforcing Long-term User Engagement.

Download
3.99 MB

MP4 File (1148-2min-promo.mp4)

Presentation video for PrefRec: Recommender Systems with Human Preferences for Reinforcing Long-term User Engagement.

Download
3.99 MB

MP4 File (1148-2min-promo.mp4)

Presentation video for PrefRec: Recommender Systems with Human Preferences for Reinforcing Long-term User Engagement.

Download
3.99 MB

References

[1]

Gediminas Adomavicius and YoungOk Kwon. 2011. Maximizing aggregate recommendation diversity: A graph-theoretic approach. In Proc. of the 1st International Workshop on Novelty and Diversity in Recommender Systems. 3--10.

Abstract

Supplementary Material

References

Cited By

Index Terms

Recommendations

Reinforcement Learning to Optimize Long-term User Engagement in Recommender Systems

Acquiring User Information Needs for Recommender Systems

User Personality and User Satisfaction with Recommender Systems

Comments

Information

Published In

Sponsors

Publisher

Publication History

Check for updates

Author Tags

Qualifiers

Conference

Acceptance Rates

Upcoming Conference

Contributors

Other Metrics

Bibliometrics

Article Metrics

Other Metrics

Citations

Cited By

View options

PDF

eReader

Login options

Full Access

Share

Share this Publication link

Share on social media

Affiliations