Bridge 2D-3D: Uncertainty-aware Hierarchical Registration Network with Domain Alignment

Authors

  • Zhixin Cheng University of Science and Technology of China
  • Jiacheng Deng University of Science and Technology of China
  • Xinjun Li University of Science and Technology of China
  • Baoqun Yin University of Science and Technology of China
  • Tianzhu Zhang University of Science and Technology of China

DOI:

https://doi.org/10.1609/aaai.v39i3.32251

Abstract

The method for image-to-point cloud registration typically determines the rigid transformation using a coarse-to-fine pipeline. However, directly and uniformly matching image patches with point cloud patches may lead to focusing on incorrect noise patches during matching while ignoring key ones. Moreover, due to the significant differences between image and point cloud modalities, it may be challenging to bridge the domain gap without specific improvements in design. To address the above issues, we innovatively propose the Uncertainty-aware Hierarchical Matching Module (UHMM) and the Adversarial Modal Alignment Module (AMAM). Within the UHMM, we model the uncertainty of critical information in image patches and facilitate multi-level fusion interactions between image and point cloud features. In the AMAM, we design an adversarial approach to reduce the domain gap between image and point cloud. Extensive experiments and ablation studies on RGB-D Scene V2 and 7-Scenes benchmarks demonstrate the superiority of our method, making it a state-of-the-art approach for image-to-point cloud registration tasks.

Published

2025-04-11

How to Cite

Cheng, Z., Deng, J., Li, X., Yin, B., & Zhang, T. (2025). Bridge 2D-3D: Uncertainty-aware Hierarchical Registration Network with Domain Alignment. Proceedings of the AAAI Conference on Artificial Intelligence, 39(3), 2491-2499. https://doi.org/10.1609/aaai.v39i3.32251

Issue

Section

AAAI Technical Track on Computer Vision II