Automatic Term Mismatch Diagnosis for Selective Query Expansion

被引:0
|
作者
Zhao, Le [1 ]
Callan, Jamie [1 ]
机构
[1] Carnegie Mellon Univ, Language Technol Inst, Pittsburgh, PA 15213 USA
关键词
Query term diagnosis; term mismatch; term expansion; Boolean conjunctive normal form queries; simulated user interactions;
D O I
暂无
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
People are seldom aware that their search queries frequently mismatch a majority of the relevant documents. This may not be a big problem for topics with a large and diverse set of relevant documents, but would largely increase the chance of search failure for less popular search needs. We aim to address the mismatch problem by developing accurate and simple queries that require minimal effort to construct. This is achieved by targeting retrieval interventions at the query terms that are likely to mismatch relevant documents. For a given topic, the proportion of relevant documents that do not contain a term measures the probability for the term to mismatch relevant documents, or the term mismatch probability. Recent research demonstrates that this probability can be estimated reliably prior to retrieval. Typically, it is used in probabilistic retrieval models to provide query dependent term weights. This paper develops a new use: Automatic diagnosis of term mismatch. A search engine can use the diagnosis to suggest manual query reformulation, guide interactive query expansion, guide automatic query expansion, or motivate other responses. The research described here uses the diagnosis to guide interactive query expansion, and create Boolean conjunctive normal form (CNF) structured queries that selectively expand 'problem' query terms while leaving the rest of the query untouched. Experiments with TREC Ad-hoc and Legal Track datasets demonstrate that with high quality manual expansion, this diagnostic approach can reduce user effort by 33%, and produce simple and effective structured queries that surpass their bag of word counterparts.
引用
收藏
页码:515 / 524
页数:10
相关论文
共 50 条
  • [1] A Framework for Automatic Query Expansion
    Imran, Hazra
    Sharan, Aditi
    WEB INFORMATION SYSTEMS AND MINING, 2010, 6318 : 386 - +
  • [2] Collaborative learning of term-based concepts for automatic query expansion
    Klink, S
    Hust, A
    Junker, M
    Dengel, A
    MACHINE LEARNING: ECML 2002, 2002, 2430 : 195 - 206
  • [3] Exploiting Underrepresented Query Aspects for Automatic Query Expansion
    Crabtree, Daniel
    Andreae, Peter
    Gao, Xiaoying
    KDD-2007 PROCEEDINGS OF THE THIRTEENTH ACM SIGKDD INTERNATIONAL CONFERENCE ON KNOWLEDGE DISCOVERY AND DATA MINING, 2007, : 191 - 200
  • [4] A New Approach for Automatic Query Expansion
    Hmeidi, Ismail
    Al-Badarneh, Amer
    Al-Qtaish, Ahmad A.
    BUSINESS TRANSFORMATION THROUGH INNOVATION AND KNOWLEDGE MANAGEMENT: AN ACADEMIC PERSPECTIVE, VOLS 3 AND 4, 2010, : 1975 - 1989
  • [5] Query difficulty, robustness, and selective application of query expansion
    Amati, G
    Carpineto, C
    Romano, G
    ADVANCES IN INFORMATION RETRIEVAL, PROCEEDINGS, 2004, 2997 : 127 - 137
  • [6] ON TERM SELECTION FOR QUERY EXPANSION
    ROBERTSON, SE
    JOURNAL OF DOCUMENTATION, 1990, 46 (04) : 359 - 364
  • [7] Mining term association rules for automatic global query expansion: Methodology and preliminary results
    Wei, J
    Bressan, S
    Ooi, BC
    PROCEEDINGS OF THE FIRST INTERNATIONAL CONFERENCE ON WEB INFORMATION SYSTEMS ENGINEERING, VOL I, 2000, : 366 - 373
  • [8] LOCAL FEEDBACK AND INTELLIGENT AUTOMATIC QUERY EXPANSION
    PIETILAINEN, P
    INFORMATION PROCESSING & MANAGEMENT, 1983, 19 (01) : 51 - 58
  • [9] Towards the development of heuristics for automatic query expansion
    Vilares, J
    Vilares, M
    Alonso, MA
    DATABASE AND EXPERT SYSTEMS APPLICATIONS, 2001, 2113 : 887 - 896
  • [10] On the number of terms used in automatic query expansion
    Paul Ogilvie
    Ellen Voorhees
    Jamie Callan
    Information Retrieval, 2009, 12 : 666 - 679