Offline Meta-level Model-based Reinforcement Learning Approach for Cold-Start Recommendation

12/04/2020
by   Yanan Wang, et al.
0

Reinforcement learning (RL) has shown great promise in optimizing long-term user interest in recommender systems. However, existing RL-based recommendation methods need a large number of interactions for each user to learn a robust recommendation policy. The challenge becomes more critical when recommending to new users who have a limited number of interactions. To that end, in this paper, we address the cold-start challenge in the RL-based recommender systems by proposing a meta-level model-based reinforcement learning approach for fast user adaptation. In our approach, we learn to infer each user's preference with a user context variable that enables recommendation systems to better adapt to new users with few interactions. To improve adaptation efficiency, we learn to recover the user policy and reward from only a few interactions via an inverse reinforcement learning method to assist a meta-level recommendation agent. Moreover, we model the interaction relationship between the user model and recommendation agent from an information-theoretic perspective. Empirical results show the effectiveness of the proposed method when adapting to new users with only a single interaction sequence. We further provide a theoretical analysis of the recommendation performance bound.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
08/10/2022

Deep Reinforcement Learning for Dynamic Recommendation with Model-agnostic Counterfactual Policy Synthesis

Recent advances in recommender systems have proved the potential of Rein...
research
05/24/2022

Meta Policy Learning for Cold-Start Conversational Recommendation

Conversational recommender systems (CRS) explicitly solicit users' prefe...
research
06/01/2021

Improving Long-Term Metrics in Recommendation Systems using Short-Horizon Offline RL

We study session-based recommendation scenarios where we want to recomme...
research
03/11/2023

User Retention-oriented Recommendation with Decision Transformer

Improving user retention with reinforcement learning (RL) has attracted ...
research
01/19/2022

Online POI Recommendation: Learning Dynamic Geo-Human Interactions in Streams

In this paper, we focus on the problem of modeling dynamic geo-human int...
research
08/26/2023

A Comparative Study on Reward Models for UI Adaptation with Reinforcement Learning

Adapting the User Interface (UI) of software systems to user requirement...
research
08/01/2019

Reinforcement Learning for Personalized Dialogue Management

Language systems have been of great interest to the research community a...

Please sign up or login with your details

Forgot password? Click here to reset