摘要
This research proposes FFDCFormer, a novel framework for depth completion that reconstructs dense depth maps from sparse or incomplete inputs guided by color images, while maintaining both local detail preservation and global structural consistency. FFDCFormer is designed as a progressive feature-fusion encoder-decoder architecture that couples neighborhood attention-based vision Transformer with convolutional attention mechanisms, enabling joint preservation of local details and global structure. By hierarchically integrating multi-scale features, it captures fine-grained local structures and gradually aggregates global semantics. A depth refinement network is then appended to enhance cross-region geometric consistency. Experiments on the NYUv2 dataset demonstrate that FFDCFormer achieves state-of-the-art performance across multiple evaluation metrics. Furthermore, an occlusion study is conducted as a new evaluation paradigm, confirming the robustness of FFDCFormer under structured sparsity and underscoring its effectiveness and practicality in real-world scenarios.
| 源语言 | 英语 |
|---|---|
| 页(从-至) | 2474-2479 |
| 页数 | 6 |
| 期刊 | Youth Academic Annual Conference of Chinese Association of Automation, YAC |
| 期 | 2026 |
| DOI | |
| 出版状态 | 已出版 - 2026 |
| 已对外发布 | 是 |
| 活动 | 41st Youth Academic Annual Conference of Chinese Association of Automation, YAC 2026 - Changsha, 中国 期限: 8 5月 2026 → 10 5月 2026 |
学术指纹
探究 'FFDCFormer: A Progressive Feature Fusion Transformer Framework for Depth Completion' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver