Align while search: Belief- guided exploratory inference for world-grounded embodied agents, 2025

Seohui Bae, Jeonghye Kim, Youngchul Sung, Woohyung Lim · 2025 · arXiv 2512.24461

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

representative citing papers

ECHO: Learning Epistemically Adaptive Language Agents with Turn-Level Credit

cs.MA · 2026-06-29 · unverdicted · novelty 7.0

ECHO is a clipped policy-gradient method that uses posterior-sensitive rewards to give turn-level epistemic credit in multi-turn information-seeking tasks, outperforming trajectory-level GRPO on a new Clue Selector Game benchmark.

citing papers explorer

Showing 1 of 1 citing paper after filters.

ECHO: Learning Epistemically Adaptive Language Agents with Turn-Level Credit cs.MA · 2026-06-29 · unverdicted · none · ref 5
ECHO is a clipped policy-gradient method that uses posterior-sensitive rewards to give turn-level epistemic credit in multi-turn information-seeking tasks, outperforming trajectory-level GRPO on a new Clue Selector Game benchmark.

Align while search: Belief- guided exploratory inference for world-grounded embodied agents, 2025

fields

years

verdicts

representative citing papers

citing papers explorer