Clarify or answer: Reinforcement learning for agentic vqa with context under-specification.arXiv preprint arXiv:2601.16400, 2026

Zongwan Cao, Bingbing Wen, Lucy Lu Wang · 2026 · arXiv 2601.16400

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

representative citing papers

Agentic Abstention: Do Agents Know When to Stop Instead of Act?

cs.AI · 2026-06-27 · unverdicted · novelty 7.0

LLM agents often fail to abstain at the right time in uncertain multi-turn tasks, and the CONVOLVE context engineering method raises timely abstention rates on WebShop from 26.7 to 57.4 without parameter updates.

citing papers explorer

Showing 1 of 1 citing paper.

Agentic Abstention: Do Agents Know When to Stop Instead of Act? cs.AI · 2026-06-27 · unverdicted · none · ref 16
LLM agents often fail to abstain at the right time in uncertain multi-turn tasks, and the CONVOLVE context engineering method raises timely abstention rates on WebShop from 26.7 to 57.4 without parameter updates.

Clarify or answer: Reinforcement learning for agentic vqa with context under-specification.arXiv preprint arXiv:2601.16400, 2026

fields

years

verdicts

representative citing papers

citing papers explorer