A mask-conditioned encoder-decoder network predicts pixel-level depth of a scene with the masked object removed from a single RGB image, outperforming depth-filling and image-inpainting baselines.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2019 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Counterfactual Depth from a Single RGB Image
A mask-conditioned encoder-decoder network predicts pixel-level depth of a scene with the masked object removed from a single RGB image, outperforming depth-filling and image-inpainting baselines.