What makes a query difficult?
Author(s) -
David Carmel,
Elad YomTov,
Adam Darlow,
Dan Pelleg
Publication year - 2006
Publication title -
citeseer x (the pennsylvania state university)
Language(s) - English
Resource type - Conference proceedings
ISBN - 1-59593-369-7
DOI - 10.1145/1148170.1148238
Subject(s) - computer science , information retrieval , component (thermodynamics) , set (abstract data type) , web search query , query expansion , query language , web query classification , query optimization , domain (mathematical analysis) , topic model , data mining , search engine , mathematical analysis , physics , mathematics , thermodynamics , programming language
This work tries to answer the question of what makes a query difficult. It addresses a novel model that captures the main components of a topic and the relationship between those components and topic difficulty. The three components of a topic are the textual expression describing the information need (the query or queries), the set of documents relevant to the topic (the Qrels), and the entire collection of documents. We show experimentally that topic difficulty strongly depends on the distances between these components. In the absence of knowledge about one of the model components, the model is still useful by approximating the missing component based on the other components. We demonstrate the applicability of the difficulty model for several uses such as predicting query difficulty, predicting the number of topic aspects expected to be covered by the search results, and analyzing the findability of a specific domain.
Accelerating Research
Robert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom
Address
John Eccles HouseRobert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom