A diffusion-based framework improves spatial and temporal consistency of clothes in virtual try-on videos, reporting the best FID/KID and several video metrics on four public datasets.
Rerender a video: Zero-shot text-guided video-to-video translation,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
extension 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
extension 1polarities
extend 1representative citing papers
citing papers explorer
-
RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency
A diffusion-based framework improves spatial and temporal consistency of clothes in virtual try-on videos, reporting the best FID/KID and several video metrics on four public datasets.