跳到主要导航 跳到搜索 跳到主要内容

Topology-aware multi-information fusion for object recognition

  • Yuhao Wang
  • , Yong Zuo*
  • , Yi Tang*
  • , Xiaobin Hong
  • , Jian Wu
  • , Ziyu Bian
  • *此作品的通讯作者
  • Beijing University of Posts and Telecommunications
  • Beijing Institute of Technology

科研成果: 期刊稿件文章同行评审

摘要

Multi-source information fusion plays a crucial role in enhancing object recognition performance. However, in real-world applications such as autonomous driving and industrial inspection, occlusion and data inconsistencies caused by complex environments can undermine the reliability of extracted features. In this paper, we propose a Topology-Aware Multi-Information Fusion (TMF) model, designed to improve the robustness and generalizability of feature extraction in multi-sensor data. For the first time, our model simultaneously integrates topological architectures into both feature extraction and propagation within a multi-source information framework. The proposed model introduces two key modules. The Enhancing Feature Module (EFM) refines local geometric structures in a topology-preserving manner. The Attention Topology Module (ATM) applies topology-aware attention during feature propagation to dynamically recalibrate feature importance, thereby improving the cross-modal fusion process. In addition, 2D RGB features extracted by a lightweight convolutional encoder are concatenated with 3D point-cloud features, providing a clear and effective fusion strategy. Through a structured fusion framework, our method effectively integrates 3D point cloud features with 2D image-based convolutional descriptors, maximizing the complementary advantages of different sensor modalities. Extensive experiments conducted on the S3DIS and Semantic3D datasets validate the effectiveness of our model. Compared to the typical object recognition model, PointNet, our proposed method achieves a 15.3% and 17.1% improvement in mIoU, respectively. Additionally, we further validate the model using self-collected real-world data, demonstrating its applicability across different data distributions and its potential for real-world multi-modal object recognition.

源语言英语
文章编号18393
期刊Scientific Reports
16
1
DOI
出版状态已出版 - 12月 2026
已对外发布

指纹

探究 'Topology-aware multi-information fusion for object recognition' 的科研主题。它们共同构成独一无二的指纹。

引用此