Reinforcing User Retention in a Billion Scale Short Video Recommender System

AI-generated keywords: User Retention Short Video Platforms Reinforcement Learning Daily Active Users RLUR

AI-generated Key Points

  • Authors address the challenge of optimizing user retention on short video platforms
  • Traditional models struggle to effectively optimize retention due to its long-term nature and complexity
  • Proposed novel method called RLUR based on reinforcement learning techniques
  • Formulated problem as an infinite-horizon request-based Markov Decision Process with objective of minimizing time interval between user sessions
  • RLUR shows promising results in both offline and live experiments, implemented in Kuaishou app
  • Method consistently improves user retention and DAU metrics
Also access our AI generated: Comprehensive summary, Lay summary, Blog-like article; or ask questions about this paper to our AI assistant.

Authors: Qingpeng Cai, Shuchang Liu, Xueliang Wang, Tianyou Zuo, Wentao Xie, Bin Yang, Dong Zheng, Peng Jiang, Kun Gai

The Web Conference 2023 Industry Track
License: CC ZERO 1.0

Abstract: Recently, short video platforms have achieved rapid user growth by recommending interesting content to users. The objective of the recommendation is to optimize user retention, thereby driving the growth of DAU (Daily Active Users). Retention is a long-term feedback after multiple interactions of users and the system, and it is hard to decompose retention reward to each item or a list of items. Thus traditional point-wise and list-wise models are not able to optimize retention. In this paper, we choose reinforcement learning methods to optimize the retention as they are designed to maximize the long-term performance. We formulate the problem as an infinite-horizon request-based Markov Decision Process, and our objective is to minimize the accumulated time interval of multiple sessions, which is equal to improving the app open frequency and user retention. However, current reinforcement learning algorithms can not be directly applied in this setting due to uncertainty, bias, and long delay time incurred by the properties of user retention. We propose a novel method, dubbed RLUR, to address the aforementioned challenges. Both offline and live experiments show that RLUR can significantly improve user retention. RLUR has been fully launched in Kuaishou app for a long time, and achieves consistent performance improvement on user retention and DAU.

Submitted to arXiv on 03 Feb. 2023

Ask questions about this paper to our AI assistant

You can also chat with multiple papers at once here.

AI assistant instructions?

Results of the summarizing process for the arXiv paper: 2302.01724v1

In their paper titled "Reinforcing User Retention in a Billion Scale Short Video Recommender System," authors Qingpeng Cai, Shuchang Liu, Xueliang Wang, Tianyou Zuo, Wentao Xie, Bin Yang, Dong Zheng, Peng Jiang, and Kun Gai address the challenge of optimizing user retention on short video platforms. These platforms have experienced rapid growth by recommending engaging content to users with the aim of increasing Daily Active Users (DAU). However, traditional models struggle to effectively optimize retention due to its long-term nature and complexity. To tackle this issue, the authors propose a novel method called RLUR based on reinforcement learning techniques. The authors formulate the problem as an infinite-horizon request-based Markov Decision Process with the objective of minimizing the time interval between multiple user sessions to improve app open frequency and user retention. Despite challenges such as uncertainty and bias in current reinforcement learning algorithms caused by user retention properties, RLUR shows promising results in both offline and live experiments. The method has been successfully implemented in the Kuaishou app and consistently improves user retention and DAU metrics. Overall, this research contributes valuable insights into enhancing user retention strategies on short video platforms through innovative reinforcement learning approaches.
Created on 22 Sep. 2026

Assess the quality of the AI-generated content by voting

Score: 0

Why do we need votes?

Votes are used to determine whether we need to re-run our summarizing tools. If the count reaches -10, our tools can be restarted.

Similar papers summarized with our AI tools

Navigate through even more similar papers through a

tree representation

Look for similar papers (in beta version)

By clicking on the button above, our algorithm will scan all papers in our database to find the closest based on the contents of the full papers and not just on metadata. Please note that it only works for papers that we have generated summaries for and you can rerun it from time to time to get a more accurate result while our database grows.

Disclaimer: The AI-based summarization tool and virtual assistant provided on this website may not always provide accurate and complete summaries or responses. We encourage you to carefully review and evaluate the generated content to ensure its quality and relevance to your needs.