GPT-3.5 and Llama-3-70B recall about 59-60% of long-tail historical entities versus ReLiK's 45.7%, but their lower precision leaves F1 scores near 53 versus ReLiK's 56.1.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Evaluation of LLMs on Long-tail Entity Linking in Historical Documents
GPT-3.5 and Llama-3-70B recall about 59-60% of long-tail historical entities versus ReLiK's 45.7%, but their lower precision leaves F1 scores near 53 versus ReLiK's 56.1.