PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
August 2, 20243 citationsOpen Access

Query Variability and Experimental Consistency: A Concerning Case Study

View Full Paper
LRLida RashidiThe University of MelbourneJZJustin ZobelThe University of MelbourneAMAlistair MoffatThe University of Melbourne

Key Points

Key points are not available for this paper at this time.

Abstract

In offline experimentation, the effectiveness of a search engine is evaluated using a document collection, a set of queries against that collection, a set of relevance judgments connecting the documents and the queries, and an effectiveness metric. This measurement pipeline is used as a surrogate for user satisfaction - the extent to which the system provides useful information to the users that are issuing the queries. But queries are responses to information needs, or topics, and there can be a wide variety of ways in which any given information need can be expressed as a query. That one-to-many relationship suggests that, in an IR experiment, use of any single query to represent a topic may be insufficient. In this case study, we demonstrate that this practice is indeed a weakness, by showing that the TREC 2013 and 2014 Web track queries, which are regarded as being indicative of specific information needs, are not necessarily representative of crowd-generated queries for the same underlying needs, and can give rise to inconsistent system relativities when compared to user-generated queries. From this instance we must thus note an element of concern: that current test collection design strategies can lead to effectiveness results that are at odds with those experienced by typical non-expert users.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Rashidi et al. (2024) studied this question.

synapsesocial.com/papers/68e5dae2b6db643587570670https://doi.org/10.1145/3664190.3672519
Ask AI
Helpful
Bookmark
Share
View Full Paper