A Novel Monocular Disparity Estimation Network with Domain Transformation and Ambiguity Learning

by   Juan Luis Gonzalez Bello, et al.

Convolutional neural networks (CNN) have shown state-of-the-art results for low-level computer vision problems such as stereo and monocular disparity estimations, but still, have much room to further improve their performance in terms of accuracy, numbers of parameters, etc. Recent works have uncovered the advantages of using an unsupervised scheme to train CNN's to estimate monocular disparity, where only the relatively-easy-to-obtain stereo images are needed for training. We propose a novel encoder-decoder architecture that outperforms previous unsupervised monocular depth estimation networks by (i) taking into account ambiguities, (ii) efficient fusion between encoder and decoder features with rectangular convolutions and (iii) domain transformations between encoder and decoder. Our architecture outperforms the Monodepth baseline in all metrics, even with a considerable reduction of parameters. Furthermore, our architecture is capable of estimating a full disparity map in a single forward pass, whereas the baseline needs two passes. We perform extensive experiments to verify the effectiveness of our method on the KITTI dataset.


page 1

page 3

page 4


Finding Correspondences for Optical Flow and Disparity Estimations using a Sub-pixel Convolution-based Encoder-Decoder Network

Deep convolutional neural networks (DCNN) have recently shown promising ...

Learning Monocular Depth Estimation via Selective Distillation of Stereo Knowledge

Monocular depth estimation has been extensively explored based on deep l...

Deep 3D-Zoom Net: Unsupervised Learning of Photo-Realistic 3D-Zoom

The 3D-zoom operation is the positive translation of the camera in the Z...

MEStereo-Du2CNN: A Novel Dual Channel CNN for Learning Robust Depth Estimates from Multi-exposure Stereo Images for HDR 3D Applications

Display technologies have evolved over the years. It is critical to deve...

Deep 3D Pan via adaptive "t-shaped" convolutions with global and local adaptive dilations

Recent advances in deep learning have shown promising results in many lo...

Lightweight Monocular Depth Estimation through Guided Decoding

We present a lightweight encoder-decoder architecture for monocular dept...

Please sign up or login with your details

Forgot password? Click here to reset