Pith. sign in

A simple recipe for contrastively pre-training video-first encoders beyond 16 frames.2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 14386–14397,

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CV 1

years

2026 1

verdicts

CONDITIONAL 1

representative citing papers

Gen4U: Unifying Video Generation and Understanding via Diffusion

cs.CV · 2026-07-07 · conditional · novelty 6.0

Frozen video diffusion models, probed at optimal depth and noise levels, produce representations competitive with discriminative encoders across semantic and geometric video tasks in a single forward pass.

citing papers explorer

Showing 1 of 1 citing paper.

  • Gen4U: Unifying Video Generation and Understanding via Diffusion cs.CV · 2026-07-07 · conditional · none · ref 12

    Frozen video diffusion models, probed at optimal depth and noise levels, produce representations competitive with discriminative encoders across semantic and geometric video tasks in a single forward pass.