A temporal diffusion pretraining stage aligned with frame latents improves self-supervised monocular depth estimation in endoscopic video.
Unsupervised learning of depth and ego-motion from video
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
eess.IV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
MetaFE-DE: Learning Meta Feature Embedding for Depth Estimation from Monocular Endoscopic Images
A temporal diffusion pretraining stage aligned with frame latents improves self-supervised monocular depth estimation in endoscopic video.