A TREC benchmark track found that LLM-prompted ranking systems beat the previous four-time winner, and that synthetic queries from T5 and GPT-4 rank systems nearly as consistently as human queries (Kendall tau = 0.8487).
Significant improvements over the state of the art? a case study of the ms marco document ranking leaderboard
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.IR 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Overview of the TREC 2023 deep learning track
A TREC benchmark track found that LLM-prompted ranking systems beat the previous four-time winner, and that synthetic queries from T5 and GPT-4 rank systems nearly as consistently as human queries (Kendall tau = 0.8487).