A temporal diffusion pretraining stage aligned with frame latents improves self-supervised monocular depth estimation in endoscopic video.
Monocular real- time hand shape and motion capture using multi-modal data
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
eess.IV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
MetaFE-DE: Learning Meta Feature Embedding for Depth Estimation from Monocular Endoscopic Images
A temporal diffusion pretraining stage aligned with frame latents improves self-supervised monocular depth estimation in endoscopic video.