Pith. sign in

Args: Alignment as reward-guided search

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

Controlling Multimodal LLMs via Reward-guided Decoding

cs.CV · 2025-08-15 · conditional · novelty 6.0

MRGD guides MLLM decoding with a learned hallucination reward and a detector-based recall reward, allowing users to trade off object precision, recall, and test-time compute while reducing object hallucinations on CHAIR and AMBER.

citing papers explorer

Showing 1 of 1 citing paper.

  • Controlling Multimodal LLMs via Reward-guided Decoding cs.CV · 2025-08-15 · conditional · none · ref 19

    MRGD guides MLLM decoding with a learned hallucination reward and a detector-based recall reward, allowing users to trade off object precision, recall, and test-time compute while reducing object hallucinations on CHAIR and AMBER.