Learning depth from single monocular images using deep convolutional neural fields

Liu, F.; Shen, C.; Lin, G.; Reid, I.

Please use this identifier to cite or link to this item: https://hdl.handle.net/2440/106735

Scopus	Web of Science®	Altmetric
Citations
?	?

Full metadata record

DC Field	Value	Language
dc.contributor.author	Liu, F.	-
dc.contributor.author	Shen, C.	-
dc.contributor.author	Lin, G.	-
dc.contributor.author	Reid, I.	-
dc.date.issued	2016	-
dc.identifier.citation	IEEE Transactions on Pattern Analysis and Machine Intelligence, 2016; 38(10):2024-2039	-
dc.identifier.issn	0162-8828	-
dc.identifier.issn	2160-9292	-
dc.identifier.uri	http://hdl.handle.net/2440/106735	-
dc.description	Date of publication 2 Dec. 2015; date of current version 12 Sept. 2016.	-
dc.description.abstract	In this article, we tackle the problem of depth estimation from single monocular images. Compared with depth estimation using multiple images such as stereo depth perception, depth from monocular images is much more challenging. Prior work typically focuses on exploiting geometric priors or additional sources of information, most using hand-crafted features. Recently, there is mounting evidence that features from deep convolutional neural networks (CNN) set new records for various vision applications. On the other hand, considering the continuous characteristic of the depth values, depth estimation can be naturally formulated as a continuous conditional random field (CRF) learning problem. Therefore, here we present a deep convolutional neural field model for estimating depths from single monocular images, aiming to jointly explore the capacity of deep CNN and continuous CRF. In particular, we propose a deep structured learning scheme which learns the unary and pairwise potentials of continuous CRF in a unified deep CNN framework. We then further propose an equally effective model based on fully convolutional networks and a novel superpixel pooling method, which is about 10 times faster, to speedup the patch-wise convolutions in the deep model. With this more efficient model, we are able to design deeper networks to pursue better performance. Our proposed method can be used for depth estimation of general scenes with no geometric priors nor any extra information injected. In our case, the integral of the partition function can be calculated in a closed form such that we can exactly solve the log-likelihood maximization. Moreover, solving the inference problem for predicting depths of a test image is highly efficient as closed-form solutions exist. Experiments on both indoor and outdoor scene datasets demonstrate that the proposed method outperforms state-of-the-art depth estimation approaches.	-
dc.description.statementofresponsibility	Fayao Liu, Chunhua Shen, Guosheng Lin, and Ian Reid	-
dc.language.iso	en	-
dc.publisher	IEEE	-
dc.rights	© 2015 IEEE	-
dc.source.uri	http://dx.doi.org/10.1109/tpami.2015.2505283	-
dc.subject	Depth estimation; conditional random field (CRF); deep convolutional neural networks (CNN); fully convolutional networks; superpixel pooling	-
dc.title	Learning depth from single monocular images using deep convolutional neural fields	-
dc.type	Journal article	-
dc.identifier.doi	10.1109/TPAMI.2015.2505283	-
pubs.publication-status	Published	-
dc.identifier.orcid	Reid, I. [0000-0001-7790-6423]	-
Appears in Collections:	Aurora harvest 8 Computer Science publications

Files in This Item:

File	Description	Size	Format
RA_hdl_106735.pdf Restricted Access	Restricted Access	1.88 MB	Adobe PDF	View/Open

Show simple item record

Adelaide Research & Scholarship