Diagnostic Evaluation of Information Retrieval Models

被引：44

作者：

Fang, Hui ^{[1
]}

Tao, Tao ^{[2
]}

Zhai, Chengxiang ^{[3
]}

机构：

[1] Univ Delaware, Dept Elect & Comp Engn, Newark, DE 19716 USA

[2] Microsoft Corp, Redmond, WA 98052 USA

[3] Univ Illinois, Dept Comp Sci, Urbana, IL 61801 USA

来源：

ACM TRANSACTIONS ON INFORMATION SYSTEMS | 2011年 / 29卷 / 02期

基金：

美国国家科学基金会;

关键词：

Algorithms; Experimentation; Measurement; Retrieval heuristics; constraints; formal models; TF-IDF weighting; diagnostic evaluation; PROBABILISTIC MODELS;

D O I：

10.1145/1961209.1961210

中图分类号：

TP [自动化技术、计算机技术];

学科分类号：

0812 ;

摘要：

Developing effective retrieval models is a long-standing central challenge in information retrieval research. In order to develop more effective models, it is necessary to understand the deficiencies of the current retrieval models and the relative strengths of each of them. In this article, we propose a general methodology to analytically and experimentally diagnose the weaknesses of a retrieval function, which provides guidance on how to further improve its performance. Our methodology is motivated by the empirical observation that good retrieval performance is closely related to the use of various retrieval heuristics. We connect the weaknesses and strengths of a retrieval function with its implementations of these retrieval heuristics, and propose two strategies to check how well a retrieval function implements the desired retrieval heuristics. The first strategy is to formalize heuristics as constraints, and use constraint analysis to analytically check the implementation of retrieval heuristics. The second strategy is to define a set of relevance-preserving perturbations and perform diagnostic tests to empirically evaluate how well a retrieval function implements retrieval heuristics. Experiments show that both strategies are effective to identify the potential problems in implementations of the retrieval heuristics. The performance of retrieval functions can be improved after we fix these problems.

引用

页数：42

共 50 条

[1] Interactive Information Retrieval: Models, Algorithms, and Evaluation
Zhai, ChengXiang
[J]. PROCEEDINGS OF THE 43RD INTERNATIONAL ACM SIGIR CONFERENCE ON RESEARCH AND DEVELOPMENT IN INFORMATION RETRIEVAL (SIGIR '20), 2020, : 2444 - 2447
[2] Interactive Information Retrieval: Models, Algorithms, and Evaluation
Zhai, ChengXiang
[J]. SIGIR '21 - PROCEEDINGS OF THE 44TH INTERNATIONAL ACM SIGIR CONFERENCE ON RESEARCH AND DEVELOPMENT IN INFORMATION RETRIEVAL, 2021, : 2662 - 2665
[3] AN EVALUATION OF TERM DEPENDENCE MODELS IN INFORMATION-RETRIEVAL
SALTON, G
BUCKLEY, C
YU, CT
[J]. LECTURE NOTES IN COMPUTER SCIENCE, 1983, 146 : 151 - 173
[4] Evaluation of News Search Engines Based On Information Retrieval Models
Bokhari M.U.
Adhami M.K.
Ahmad A.
[J]. Operations Research Forum, 2 (3)
[5] On the analysis and evaluation of information retrieval models for social book search
Irfan Ullah
Shah Khusro
[J]. Multimedia Tools and Applications, 2023, 82 : 6431 - 6478
[6] On the analysis and evaluation of information retrieval models for social book search
Ullah, Irfan
Khusro, Shah
[J]. MULTIMEDIA TOOLS AND APPLICATIONS, 2023, 82 (05) : 6431 - 6478
[7] Information Retrieval Evaluation
Hartley, Dick
[J]. EDUCATION FOR INFORMATION, 2011, 28 (2-4) : 341 - 342
[8] Information retrieval models in the context of retrieval tasks
O. L. Golitsyna
N. V. Maksimov
[J]. Automatic Documentation and Mathematical Linguistics, 2011, 45 (1) : 20 - 32
[9] Information Retrieval Models in the Context of Retrieval Tasks
Golitsyna, O. L.
Maksimov, N. V.
[J]. AUTOMATIC DOCUMENTATION AND MATHEMATICAL LINGUISTICS, 2011, 45 (01) : 20 - 32
[10] Performance Evaluation of Information Retrieval Models in Bug Localization on the Method Level
Alduailij, Mai
Al-Duailej, Mona
[J]. PROCEEDINGS OF THE 2015 INTERNATIONAL CONFERENCE ON COLLABORATION TECHNOLOGIES AND SYSTEMS, 2015, : 305 - 313

← 1 2 3 4 5 →