Mask 2D-3D: Adaptive Dual-Masked Autoencoder Network for Image-to-Point Cloud Registration

Researchers proposed an adaptive dual-masked autoencoder network for image-to-point cloud registration, addressing limitations in standard masked autoencoders. The Intermodal Dual-MAE Framework (ID-MAE) uses a Similarity-based RL Masking Strategy (SRLM) to enhance cross-modal representation learning and improve 2D-3D correspondence estimation.

RSS Score 0 9/17/2026, 4:00:00 AM Original Source
Save an API key to vote.