IINet: Implicit Intra-inter Information Fusion for Real-Time Stereo Matching
DOI:
https://doi.org/10.1609/aaai.v38i4.28107Keywords:
CV: 3D Computer VisionAbstract
Recently, there has been a growing interest in 3D CNN-based stereo matching methods due to their remarkable accuracy. However, the high complexity of 3D convolution makes it challenging to strike a balance between accuracy and speed. Notably, explicit 3D volumes contain considerable redundancy. In this study, we delve into more compact 2D implicit network to eliminate redundancy and boost real-time performance. However, simply replacing explicit 3D networks with 2D implicit networks causes issues that can lead to performance degradation, including the loss of structural information, the quality decline of inter-image information, as well as the inaccurate regression caused by low-level features. To address these issues, we first integrate intra-image information to fuse with inter-image information, facilitating propagation guided by structural cues. Subsequently, we introduce the Fast Multi-scale Score Volume (FMSV) and Confidence Based Filtering (CBF) to efficiently acquire accurate multi-scale, noise-free inter-image information. Furthermore, combined with the Residual Context-aware Upsampler (RCU), our Intra-Inter Fusing network is meticulously designed to enhance information transmission on both feature-level and disparity-level, thereby enabling accurate and robust regression. Experimental results affirm the superiority of our network in terms of both speed and accuracy compared to all other fast methods.Downloads
Published
2024-03-24
How to Cite
Li, X., Zhang, C., Su, W., & Tao, W. (2024). IINet: Implicit Intra-inter Information Fusion for Real-Time Stereo Matching. Proceedings of the AAAI Conference on Artificial Intelligence, 38(4), 3225-3233. https://doi.org/10.1609/aaai.v38i4.28107
Issue
Section
AAAI Technical Track on Computer Vision III