vix.ing · top · new · best · stats · spec

Overcoming Prior Misspecification in Online Learning to Rank

2023/01/25 by Javad Azizi, Ofer Meshi, Azizi, Javad +5 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms

paper · pdf · doi:10.48550/arxiv.2301.10651

openalex publication_date 2023/01/25 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

The recent literature on online learning to rank (LTR) has established the utility of prior knowledge to Bayesian ranking bandit algorithms. However, a major limitation of existing work is the requirement for the prior used by the algorithm to match the true prior. In this paper, we propose and analyze adaptive algorithms that address this issue and additionally extend these results to the linear and generalized linear models. We also consider scalar relevance feedback on top of click feedback. Moreover, we demonstrate the efficacy of our algorithms using both synthetic and real-world experiments.

Cited by

Related