SOTAVerified

Further Theoretical Study of Distribution Separation Method for Information Retrieval

2015-10-16Unverified0· sign in to hype

Zhang Peng, Yu Qian, Hou Yuexian, Song Dawei, Li Jingfei, Hu Bin

Unverified — Be the first to reproduce this paper.

Reproduce

Abstract

Recently, a Distribution Separation Method (DSM) is proposed for relevant feedback in information retrieval, which aims to approximate the true relevance distribution by separating a seed irrelevance distribution from the mixture one. While DSM achieved a promising empirical performance, theoretical analysis of DSM is still need further study and comparison with other relative retrieval model. In this article, we first generalize DSM's theoretical property, by proving that its minimum correlation assumption is equivalent to the maximum (original and symmetrized) KL-Divergence assumption. Second, we also analytically show that the EM algorithm in a well-known Mixture Model is essentially a distribution separation process and can be simplified using the linear separation algorithm in DSM. Some empirical results are also presented to support our theoretical analysis.

Tasks

Reproductions