vix.ing · top · new · best · stats

3D LiDAR and Stereo Fusion using Stereo Matching Network with Conditional Cost Volume Normalization

2019/04/05 by Tsun-Hsuan Wang, Wang, Tsun-Hsuan, Hou-Ning Hu +11 · 2 citations
Computer Science · Engineering · #Advanced Vision and Imaging #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Processing Techniques and Applications #Photoacoustic and Ultrasonic Imaging #cs.CV

paper · pdf · doi:10.48550/arxiv.1904.02917

ver.1

arxiv created 2019/04/05 · openalex publication_date 2019/04/05 · arxiv updated 2019/04/08 · openalex created_date 2020/07/16 · openalex updated_date 2026/07/28

Abstract

The complementary characteristics of active and passive depth sensing techniques motivate the fusion of the Li-DAR sensor and stereo camera for improved depth perception. Instead of directly fusing estimated depths across LiDAR and stereo modalities, we take advantages of the stereo matching network with two enhanced techniques: Input Fusion and Conditional Cost Volume Normalization (CCVNorm) on the LiDAR information. The proposed framework is generic and closely integrated with the cost volume component that is commonly utilized in stereo matching neural networks. We experimentally verify the efficacy and robustness of our method on the KITTI Stereo and Depth Completion datasets, obtaining favorable performance against various fusion strategies. Moreover, we demonstrate that, with a hierarchical extension of CCVNorm, the proposed method brings only slight overhead to the stereo matching network in terms of computation time and model size. For project page, see https://zswang666.github.io/Stereo-LiDAR-CCVNorm-Project-Page/

Cited by

Related