vix.ing · top · new · best · stats · spec

On the Synergies between Machine Learning and Binocular Stereo for Depth Estimation from Images: a Survey

2020/04/18 by Matteo Poggi, Poggi, Matteo, Fabio Tosi +7
Computer Science · Engineering · #Advanced Image Processing Techniques #Advanced Vision and Imaging #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Processing Techniques and Applications #Optical measurement and interference techniques

paper · pdf · doi:10.48550/arxiv.2004.08566

openalex publication_date 2020/04/18 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

Stereo matching is one of the longest-standing problems in computer vision with close to 40 years of studies and research. Throughout the years the paradigm has shifted from local, pixel-level decision to various forms of discrete and continuous optimization to data-driven, learning-based methods. Recently, the rise of machine learning and the rapid proliferation of deep learning enhanced stereo matching with new exciting trends and applications unthinkable until a few years ago. Interestingly, the relationship between these two worlds is two-way. While machine, and especially deep, learning advanced the state-of-the-art in stereo matching, stereo itself enabled new ground-breaking methodologies such as self-supervised monocular depth estimation based on deep networks. In this paper, we review recent research in the field of learning-based depth estimation from single and binocular images highlighting the synergies, the successes achieved so far and the open challenges the community is going to face in the immediate future.

Related