GPT-3.5 with chain-of-thought prompting still fails often on multi-hop reasoning with external knowledge, especially with distractors, counterfactual facts, non-sequential proof structures, and higher hop counts.
Language models are few-shot learners
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Large Language Models Still Face Challenges in Multi-Hop Reasoning with External Knowledge
GPT-3.5 with chain-of-thought prompting still fails often on multi-hop reasoning with external knowledge, especially with distractors, counterfactual facts, non-sequential proof structures, and higher hop counts.