Generalized Ensemble Model for Document Ranking in Information Retrieval

被引:3
|
作者
Wang, Yanshan [1 ]
Choi, In-Chan [2 ]
Liu, Hongfang [1 ]
机构
[1] Mayo Clin, Dept Hlth Sci Res, Rochester, MN 55905 USA
[2] Korea Univ, Sch Ind Management Engn, Seoul 136701, South Korea
关键词
information retrieval; optimization; mean average precision; document ranking; ensemble model;
D O I
10.2298/CSIS160229042W
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
A generalized ensemble model (gEnM) for document ranking is proposed in this paper. The gEnM linearly combines the document retrieval models and tries to retrieve relevant documents at high positions. In order to obtain the optimal linear combination of multiple document retrieval models or rankers, an optimization program is formulated by directly maximizing the mean average precision. Both supervised and unsupervised learning algorithms are presented to solve this program. For the supervised scheme, two approaches are considered based on the data setting, namely batch and online setting. In the batch setting, we propose a revised Newton's algorithm, gEnM. BAT, by approximating the derivative and Hessian matrix. In the online setting, we advocate a stochastic gradient descent (SGD) based algorithm-gEnM. ON. As for the unsupervised scheme, an unsupervised ensemble model (UnsEnM) by iteratively co-learning from each constituent ranker is presented. Experimental study on benchmark data sets verifies the effectiveness of the proposed algorithms. Therefore, with appropriate algorithms, the gEnM is a viable option in diverse practical information retrieval applications.
引用
收藏
页码:123 / 151
页数:29
相关论文
共 50 条
  • [1] A probabilistic information retrieval model by document ranking using term dependencies
    You, Hyun-Jo
    Lee, Jung-Jin
    [J]. KOREAN JOURNAL OF APPLIED STATISTICS, 2019, 32 (05) : 763 - 782
  • [2] An Ensemble Click Model for Web Document Ranking
    Bakhtiarvand, D. Bidekani
    Farzi, S.
    [J]. INTERNATIONAL JOURNAL OF ENGINEERING, 2020, 33 (07): : 1208 - 1213
  • [3] Refining aggregation functions for improving document ranking in information retrieval
    Boughanem, Mohand
    Loiseau, Yannick
    Prade, Henri
    [J]. SCALABLE UNCERTAINTY MANAGEMENT, PROCEEDINGS, 2007, 4772 : 255 - +
  • [4] Analysis of Probabilistic model for Document Retrieval in Information Retrieval
    Tamrakar, Astha
    Vishwakarma, Santosh K.
    [J]. 2015 INTERNATIONAL CONFERENCE ON COMPUTATIONAL INTELLIGENCE AND COMMUNICATION NETWORKS (CICN), 2015, : 760 - 765
  • [5] A Novel Factoid Ranking Model for Information Retrieval
    Ni, Youcong
    Wang, Wei
    [J]. ADVANCES IN WEB AND NETWORK TECHNOLOGIES, AND INFORMATION MANAGEMENT, PROCEEDINGS, 2007, 4537 : 308 - +
  • [6] Neural ranking models for document retrieval
    Mohamed Trabelsi
    Zhiyu Chen
    Brian D. Davison
    Jeff Heflin
    [J]. Information Retrieval Journal, 2021, 24 : 400 - 444
  • [7] Neural ranking models for document retrieval
    Trabelsi, Mohamed
    Chen, Zhiyu
    Davison, Brian D.
    Heflin, Jeff
    [J]. INFORMATION RETRIEVAL JOURNAL, 2021, 24 (06): : 400 - 444
  • [8] Document re-ranking by generality in bio-medical information retrieval
    Yan, X
    Li, X
    Song, DW
    [J]. WEB INFORMATION SYSTEMS ENGINEERING - WISE 2005, 2005, 3806 : 376 - 389
  • [9] A Neural Autoencoder Approach for Document Ranking and Query Refinement in Pharmacogenomic Information Retrieval
    Pfeiffer, Jonas
    Broscheit, Samuel
    Gemulla, Rainer
    Goeschl, Mathias
    [J]. SIGBIOMED WORKSHOP ON BIOMEDICAL NATURAL LANGUAGE PROCESSING (BIONLP 2018), 2018, : 87 - 97
  • [10] Mean-Variance Analysis: A New Document Ranking Theory in Information Retrieval
    Wang, Jun
    [J]. ADVANCES IN INFORMATION RETRIEVAL, PROCEEDINGS, 2009, 5478 : 4 - 16