A Lightweight RGB-LiDAR Feature Recalibration Network for Large-Scale 3D Scene Understanding
Weifeng Zhai, Zexi TanSemantic segmentation of large-scale 3D point clouds is a fundamental task in robotic perception, semantic mapping, and urban scene understanding. Existing methods mainly rely on geometric information, which limits their ability to distinguish semantic categories with similar spatial structures. To address this issue, this paper proposes a lightweight cross-modal feature learning framework that adaptively integrates geometric coordinates and RGB color information. By exploiting the complementary characteristics of spatial structure and visual appearance during feature encoding, the proposed method enhances feature discriminability while maintaining a compact model scale. Experiments on the Semantic3D dataset show that the proposed method achieves an mIoU of 87.1%, outperforming the original RandLA-Net and several representative approaches. Additional runtime and LiDAR-only cross-dataset experiments indicate the potential of the proposed structure for online outdoor point cloud perception.