pith. sign in

From pixels to prose: A large dataset of dense image cap- tions

6 Pith papers cite this work. Polarity classification is still indexing.

6 Pith papers citing it

citation-role summary

dataset 1

citation-polarity summary

fields

cs.CV 6

roles

dataset 1

polarities

use dataset 1

clear filters

representative citing papers

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning

cs.CV · 2026-05-27 · unverdicted · novelty 5.0

VCap pairs reference captions as witnesses with visual signals as adjudicators to deliver hypergeometric-precision rewards for RL in visual captioning, enabling an 8B model to outperform SOTA on benchmarks and improve weak-to-strong generalization.

citing papers explorer

Showing 5 of 5 citing papers after filters.