ReAct-prompted GPT-3.5 and GPT-4 underperform classical task-oriented dialogue systems on task success, but humans rate them as more satisfying despite lower success.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Exploring ReAct Prompting for Task-Oriented Dialogue: Insights and Shortcomings
ReAct-prompted GPT-3.5 and GPT-4 underperform classical task-oriented dialogue systems on task success, but humans rate them as more satisfying despite lower success.