Learning diverse rankings with multi-armed bandits

2008-07-05International Conference on Machine Learning 2008Unverified0· sign in to hype

Filip Radlinski, Robert Kleinberg, Thorsten Joachims

Unverified — Be the first to reproduce this paper.

Abstract

Algorithms for learning to rank Web documents usually assume a document's relevance is independent of other documents. This leads to learned ranking functions that produce rankings with redundant results. In contrast, user studies have shown that diversity at high ranks is often preferred. We present two online learning algorithms that directly learn a diverse ranking of documents based on users' clicking behavior. We show that these algorithms minimize abandonment, or alternatively, maximize the probability that a relevant document is found in the top k positions of a ranking. Moreover, one of our algorithms asymptotically achieves optimal worst-case performance even if users' interests change.

Tasks

Diversity Learning-To-Rank Multi-Armed Bandits

Learning diverse rankings with multi-armed bandits

Abstract

Tasks

Reproductions