Shallow-judged training data (many queries, few judgments per query) produces better BERT rerankers than deep-judged data (few queries, many judgments per query) when total training instances are matched, across MS MARCO and LongEval.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.IR 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Impact of Shallow vs. Deep Relevance Judgments on BERT-based Reranking Models
Shallow-judged training data (many queries, few judgments per query) produces better BERT rerankers than deep-judged data (few queries, many judgments per query) when total training instances are matched, across MS MARCO and LongEval.