Co-Prop uses LLM-generated audio control points to split videos into consistent sound segments and propagates keyframe masks frame-by-frame with audio inserted, improving audio-visual segmentation scores.
Semi-supervised video object segmentation via learning object-aware global-local correspondence
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Collaborative Hybrid Propagator for Temporal Misalignment in Audio-Visual Segmentation
Co-Prop uses LLM-generated audio control points to split videos into consistent sound segments and propagates keyframe masks frame-by-frame with audio inserted, improving audio-visual segmentation scores.