Unbiased Learning to Rank: Online or Offline?

04/28/2020
by   Qingyao Ai, et al.
0

How to obtain an unbiased ranking model by learning to rank with biased user feedback is an important research question for IR. Existing work on unbiased learning to rank (ULTR) can be broadly categorized into two groups – the studies on unbiased learning algorithms with logged data, namely the offline unbiased learning, and the studies on unbiased parameters estimation with real-time user interactions, namely the online learning to rank. While their definitions of unbiasness are different, these two types of ULTR algorithms share the same goal – to find the best models that rank documents based on their intrinsic relevance or utility. However, most studies on offline and online unbiased learning to rank are carried in parallel without detailed comparisons on their background theories and empirical performance. In this paper, we formalize the task of unbiased learning to rank and show that existing algorithms for offline unbiased learning and online learning to rank are just the two sides of the same coin. We evaluate six state-of-the-art ULTR algorithms and find that most of them can be used in both offline settings and online environments with or without minor modifications. Further, we analyze how different offline and online learning paradigms would affect the theoretical foundation and empirical effectiveness of each algorithm on both synthetic and real search data. Our findings could provide important insights and guideline for choosing and deploying ULTR algorithms in practice.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
04/16/2018

Unbiased Learning to Rank with Unbiased Propensity Estimation

Learning to rank with biased click data is a well-known challenge. A var...
research
08/11/2021

ULTRA: An Unbiased Learning To Rank Algorithm Toolbox

Learning to rank systems has become an important aspect of our daily lif...
research
01/25/2023

Overcoming Prior Misspecification in Online Learning to Rank

The recent literature on online learning to rank (LTR) has established t...
research
04/25/2023

THUIR at WSDM Cup 2023 Task 1: Unbiased Learning to Rank

This paper introduces the approaches we have used to participate in the ...
research
01/05/2022

Reinforcement Online Learning to Rank with Unbiased Reward Shaping

Online learning to rank (OLTR) aims to learn a ranker directly from impl...
research
06/10/2019

Variance Reduction in Gradient Exploration for Online Learning to Rank

Online Learning to Rank (OL2R) algorithms learn from implicit user feedb...
research
06/15/2019

Practical User Feedback-driven Internal Search Using Online Learning to Rank

We present a system, Spoke, for creating and searching internal knowledg...

Please sign up or login with your details

Forgot password? Click here to reset