跳到主要导航 跳到搜索 跳到主要内容

DCP-Net: Learning Detail–Context Perception via Spatial-Frequency Guidance for Tiny Object Detection in Remote Sensing Images

  • Beijing Institute of Technology
  • National Key Laboratory of Science and Technology on Space-Born Intelligent Information Processing
  • Beijing Institute of Remote Sensing Information

科研成果: 期刊稿件文章同行评审

摘要

Tiny object detection in remote sensing images haslong been challenging due to the low resolution of tiny objectsand complex backgrounds. However, existing adaptive receptivefield methods mainly focus on object and contextual information,making them susceptible to interference from complex backgrounds. This interference exacerbates the imbalance in sampleassignment and hinders accurate adaptation of receptive fieldsto object scales, resulting in insufficient representation of finegrained object details. In addition, most loss functions are proneto gradient instability in scenarios involving small objects. Toaddress the aforementioned problems, this paper proposes aDCP-Net, which learns detail–context perception under spatialfrequency guidance. To enhance the representation of detailsof tiny objects, the network is the first to leverage spatialfrequency to distinguish the structural features of objects andbackground. By introducing a spatial-frequency guided dualdilated convolution, it effectively enhances local object detailsand global semantic associations while suppressing interferencefrom complex backgrounds. Based on this convolutional unit, wefurther design a cross-layer fine-grained fusion module to recoverlost object details at a lower computational cost. Furthermore, weintroduce a Gaussian classification-regression loss to address theissues of imbalanced sample allocation and unstable gradient,while applying a decoupled probabilistic metric strategy anda multi-scale enhanced decoupled head to improve localizationaccuracy. Extensive experimental results demonstrate the effectiveness of the proposed method. Specifically, DCP-Net achievesan average precision (AP) of 27.5% on the AI-TODv2 datasetand an AP50 of 78.0% on the DOTA-v1.0 dataset.

指纹

探究 'DCP-Net: Learning Detail–Context Perception via Spatial-Frequency Guidance for Tiny Object Detection in Remote Sensing Images' 的科研主题。它们共同构成独一无二的指纹。

引用此