New benchmark ICRL4AHT reveals that history-conditioned ICRL methods fail to show robust adaptation in multi-agent Overcooked-V2, underperforming random baselines on unseen teammates and layouts.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2026 1verdicts
UNVERDICTED 1representative citing papers
citing papers explorer
-
Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork
New benchmark ICRL4AHT reveals that history-conditioned ICRL methods fail to show robust adaptation in multi-agent Overcooked-V2, underperforming random baselines on unseen teammates and layouts.