摘要
Recently, the performance improvement of BEV visual detection task has benefited from the extensive use of deformable attention. Deformable attention can easily transfer the features of the image space to the BEV space through the cross-attention mechanism. Compared with the global attention mechanism, when the feature resolution of the graph is larger, the computational consumption of deformable attention will be much smaller, so it can support larger Bird's-Eye-View (BEV) feature resolution. However, there are also shortcomings such as a small receptive field and insufficient information exchange. We propose a deformable attention mechanism for dynamic reference points. This module is to accumulate the reference points of each cross-attention layer on the basis of the previous layer, thereby effectively expanding the perceptual field of BEV features for querying in the image space. Extensive experiments on the nuScenes benchmark demonstrate the effectiveness of our method.
| 源语言 | 英语 |
|---|---|
| 页(从-至) | 886-890 |
| 页数 | 5 |
| 期刊 | IEEE Journal of Radio Frequency Identification |
| 卷 | 6 |
| DOI | |
| 出版状态 | 已出版 - 2022 |
学术指纹
探究 'Application of Dynamic Deformable Attention in Bird's-Eye-View Detection' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver