方向(向量空间)
计算机科学
人工智能
模式识别(心理学)
特征提取
目标检测
图层(电子)
对象(语法)
相似性(几何)
计算机视觉
特征(语言学)
频域
遥感
图像(数学)
数学
地理
语言学
化学
哲学
几何学
有机化学
作者
Shangdong Zheng,Zebin Wu,Yang Xu,Zhihui Wei,Antonio Plaza
标识
DOI:10.1109/tgrs.2022.3200980
摘要
Object detection in remote sensing images (RSIs) poses great difficulties due to arbitrary orientations, various scales and dense location of the targets over the ground. Recent evidence suggests that encoding the orientation information is of great use for training an accurate object detector for oriented object detection (OOD). In this paper, we propose a new frequency-domain orientation learning (FDOL) module with two main components: the frequency domain feature extraction (FFE) network and an orientation enhanced self-attention layer (OES-Layer). The FFE network models the interactions among spatial locations in the frequency domain to determine the frequency of spatial features. Then, these features are fed into our OES-Layer to learn the orientation information. Moreover, the orientation weights are adopted to guide the feature selection in a self-attention architecture, using them as a control gate to emphasize the spatial responses of target instances. Considering that the original similarity weights (calculated by the self-attention algorithm) do not distinctly model the orientation variation, the considered orientation weights provide an efficient asset to emphasize the orientation of objects. Extensive experiments on the DOTA and HRSC2016 datasets demonstrate that our method achieves state-of-the-art performance among single-scale methods, while achieving competitive performance over multi-scale methods.
科研通智能强力驱动
Strongly Powered by AbleSci AI