Pith. sign in

REVIEW 1 cited by

Estimating Depth from Monocular Images as Classification Using Deep Fully Convolutional Residual Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1605.02305 v3 pith:GFBRI3JZ submitted 2016-05-08 cs.CV

classification cs.CV
keywords depthdeepnetworksclassificationconvolutionallabelresidualapproach
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Depth estimation from single monocular images is a key component of scene understanding and has benefited largely from deep convolutional neural networks (CNN) recently. In this article, we take advantage of the recent deep residual networks and propose a simple yet effective approach to this problem. We formulate depth estimation as a pixel-wise classification task. Specifically, we first discretize the continuous depth values into multiple bins and label the bins according to their depth range. Then we train fully convolutional deep residual networks to predict the depth label of each pixel. Performing discrete depth label classification instead of continuous depth value regression allows us to predict a confidence in the form of probability distribution. We further apply fully-connected conditional random fields (CRF) as a post processing step to enforce local smoothness interactions, which improves the results. We evaluate our approach on both indoor and outdoor datasets and achieve state-of-the-art performance.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Unsupervised Video Depth Estimation Based on Ego-motion and Disparity Consensus

    cs.CV 2019-09 reject novelty 3.0 of 10

    A monocular depth estimator that adds stereo left-right reconstruction and disparity consistency losses to an ego-motion-based view synthesis framework achieves modest gains on KITTI depth metrics.

Pith tools