Pith. sign in

Title resolution pending

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

EgoVLM: Policy Optimization for Egocentric Video Understanding

cs.CV · 2025-06-03 · conditional · novelty 4.0

Reinforcement learning (GRPO) on non-chain-of-thought egocentric QA data lifts Qwen2.5-VL-3B to 73.7% on EgoSchema, but the result is clouded by unverified train/test separation and a supervised baseline that scores 74.0%.

citing papers explorer

Showing 1 of 1 citing paper.

  • EgoVLM: Policy Optimization for Egocentric Video Understanding cs.CV · 2025-06-03 · conditional · none · ref 7

    Reinforcement learning (GRPO) on non-chain-of-thought egocentric QA data lifts Qwen2.5-VL-3B to 73.7% on EgoSchema, but the result is clouded by unverified train/test separation and a supervised baseline that scores 74.0%.