Generalized Ensemble Model for Document Ranking in Information Retrieval
Computer Science and Information Systems, Tome 14 (2017) no. 1.

Voir la notice de l'article provenant de la source Computer Science and Information Systems website

A generalized ensemble model (gEnM) for document ranking is proposed in this paper. The gEnM linearly combines the document retrieval models and tries to retrieve relevant documents at high positions. In order to obtain the optimal linear combination of multiple document retrieval models or rankers, an optimization program is formulated by directly maximizing the mean average precision. Both supervised and unsupervised learning algorithms are presented to solve this program. For the supervised scheme, two approaches are considered based on the data setting, namely batch and online setting. In the batch setting, we propose a revised Newton’s algorithm, gEnM.BAT, by approximating the derivative and Hessian matrix. In the online setting, we advocate a stochastic gradient descent (SGD) based algorithm—gEnM.ON. As for the unsupervised scheme, an unsupervised ensemble model (UnsEnM) by iteratively co-learning from each constituent ranker is presented. Experimental study on benchmark data sets verifies the effectiveness of the proposed algorithms. Therefore, with appropriate algorithms, the gEnM is a viable option in diverse practical information retrieval applications.
Keywords: information retrieval, optimization, mean average precision, document ranking, ensemble model
@article{CSIS_2017_14_1_a7,
     author = {Yanshan Wang and In-Chan Choi and Hongfang Liu},
     title = {Generalized {Ensemble} {Model} for {Document} {Ranking} in {Information} {Retrieval}},
     journal = {Computer Science and Information Systems},
     publisher = {mathdoc},
     volume = {14},
     number = {1},
     year = {2017},
     url = {http://geodesic.mathdoc.fr/item/CSIS_2017_14_1_a7/}
}
TY  - JOUR
AU  - Yanshan Wang
AU  - In-Chan Choi
AU  - Hongfang Liu
TI  - Generalized Ensemble Model for Document Ranking in Information Retrieval
JO  - Computer Science and Information Systems
PY  - 2017
VL  - 14
IS  - 1
PB  - mathdoc
UR  - http://geodesic.mathdoc.fr/item/CSIS_2017_14_1_a7/
ID  - CSIS_2017_14_1_a7
ER  - 
%0 Journal Article
%A Yanshan Wang
%A In-Chan Choi
%A Hongfang Liu
%T Generalized Ensemble Model for Document Ranking in Information Retrieval
%J Computer Science and Information Systems
%D 2017
%V 14
%N 1
%I mathdoc
%U http://geodesic.mathdoc.fr/item/CSIS_2017_14_1_a7/
%F CSIS_2017_14_1_a7
Yanshan Wang; In-Chan Choi; Hongfang Liu. Generalized Ensemble Model for Document Ranking in Information Retrieval. Computer Science and Information Systems, Tome 14 (2017) no. 1. http://geodesic.mathdoc.fr/item/CSIS_2017_14_1_a7/