The Surprising Effectiveness of Rankers Trained on Expanded Queries

Anand, Abhijit; V, V.; Setty, Vinay; Anand, A.

The Surprising Effectiveness of Rankers Trained on Expanded Queries

Conference paper (2024)

Authors

Abhijit Anand L3S Research Center

V. V Web Information Systems -

Vinay Setty University of Stavanger

A. Anand Web Information Systems -

Research Group

Web Information Systems () (TU Delft)

Document ranking Hard queries Qpp Query rewriting

To reference this document use:

http://resolver.tudelft.nl/uuid:87dd5dd7-d105-4d42-b5d7-ac1ede38723c

More Info

expand_more

Published Date

2024

Language

English

Reuse Rights

Other than for strictly personal use, it is not permitted to download, forward or distribute the text or part of it, without the consent of the author(s) and/or copyright holder(s), unless the work is under an open content license such as Creative Commons.

Faculty

Electrical Engineering, Mathematics and Computer Science

Department

Software Technology

Research Group

Web Information Systems

Abstract

An significant challenge in text-ranking systems is handling hard queries that form the tail end of the query distribution. Difficulty may arise due to the presence of uncommon, underspecified, or incomplete queries. In this work, we improve the ranking performance of hard or difficult queries while maintaining the performance of other queries. Firstly, we do LLM-based query enrichment for training queries using relevant documents. Next, a specialized ranker is fine-tuned only on the enriched hard queries instead of the original queries. We combine the relevance scores from the specialized ranker and the base ranker, along with a query performance score estimated for each query. Our approach departs from existing methods that usually employ a single ranker for all queries, which is biased towards easy queries, which form the majority of the query distribution. In our extensive experiments on the DL-Hard dataset, we find that a principled query performance based scoring method using base and specialized ranker offers a significant improvement of up to 48.4% on the document ranking task and up to 25% on the passage ranking task compared to the baseline performance of using original queries, even outperforming SOTA model.

Files

3626772.3657938.pdf

(pdf | 1.18 Mb)

- Embargo expired in 11-01-2025

Unknown license