Skip to content
Open access

Dual-Level Spatial–Frequency Collaborative Detector for Oriented Object Detection in Remote Sensing Images

Aug 2026 · Remote Sensing · Vol 18, pp. 2845 · 0 citations · 43 references

TL;DR

The proposed DSCDet constructs a complete spatial–frequency collaborative fusion paradigm that shares a generic wavelet-based frequency extraction mechanism and cross-feature fusion module, which is adaptively deployed at both image-level and instance-level granularities.

Abstract

Oriented object detection (OOD) in remote sensing images (RSIs) suffers from insufficient feature representation caused by arbitrary rotation angles and small spatial resolutions. Existing spatial–frequency fusion paradigms merely implement single-granularity feature interaction, either global image-level frequency compensation or local instance-level feature refinement, and fail to simultaneously capture global scene semantic consistency and local object fine-grained discriminability. In this paper, we propose a unified dual-level spatial–frequency collaborative detector (DSCDet) for remote sensing OOD tasks. Different from previous decoupled designs, the proposed DSCDet constructs a complete spatial–frequency collaborative fusion paradigm that shares a generic wavelet-based frequency extraction mechanism and cross-feature fusion module, which is adaptively deployed at both image-level and instance-level granularities. Specifically, our method introduces Haar wavelet transform to extract multi-scale frequency mutation features. On this basis, a generic cross-domain attention fusion (GCDAF) is constructed with granularity-dependent positional encoding constraints. The core difference between dual granularity fusion lies in geometric positional encoding, where image-level fusion adopts global scene positional embedding to maintain overall semantic stability, and instance-level fusion leverages local pairwise instance positional embedding to optimize fine-grained target feature interaction. The unified dual-level fusion architecture comprehensively integrates global semantic integrity and local target specificity, forming a robust and universal spatial–frequency feature representation system. Extensive experiments on three public remote sensing datasets, including DOTA-v1.0, DOTA-v1.5 and DIOR-R, demonstrate that the proposed DSCDet achieves competitive and superior performance against state-of-the-art OOD detectors.

Read PDF

Similar papers

Open access 2026

MWAE-YOLO: Frequency–Spatial Collaborative Enhancement for Small-Object Detection in Remote Sensing Images

Small-objectdetection in remote sensing images remains challenging due to insufficient feature representation, weak texture information, complex backgrounds, and high sensitivity to localization errors. To address these issues, this article proposes a frequency–spatial collaborative enhancement detector, multilevel wav...

Wenqing Wang, Ding-Zhou Zhu, Han-Qing Liu · 0 citations
Open access 2026

Orientation-Aware Feature Fusion for Accurate Rotated Object Detection in Remote Sensing Images

Conventional feature fusion mechanisms largely overlook orientation information, making it difficult to effectively represent objects with diverse rotational patterns. To address this issue, we propose YOLO-RSL, a lightweight rotated object detector that introduces orientation awareness into feature representation, fea...

Jing Zhang, Mas Rina Binti Mustaffa, F. Khalid et al. · 0 citations
Open access Sep 2026

A Frequency-Spatial Segmentation Network for High-Resolution Remote Sensing Images

Semantic segmentation of high-resolution remote sensing images faces three major challenges in frequency-spatial feature fusion: background clutter mixed into high-frequency components, semantic discontinuities within large homogeneous regions, and loss of fine rigid boundaries caused by convolutional downsampling. Tra...

Qi-Yuan Zhang, Jian-Shun Liu · 0 citations
Sep 2026

Aerial Small Object Detection Based on Adaptive Dual‐Branch Frequency‐Spatial Feature Fusion

Small object detection in UAV imagery remains challenging. Existing methods still exhibit insufficient feature extraction and feature fusion capabilities, limiting their robustness under dense object distributions, complex backgrounds, blurred texture details, and scale variations. To address these issues, this paper p...

Chao Zhang, Jing-Rui Zhang, Xin Fang et al. · 0 citations
Open access 2026

High-Resolution Remote Sensing Image Segmentation Based on Spatial-Frequency Feature Enhancement and Dual Attention Fusion

High-resolution remote sensing image segmentation is a core task in remote sensing interpretation, which faces challenges such as complex distribution of ground objects, significant scale differences and blurred edges. Existing methods are relatively single, mostly focusing only on spatial feature extraction, and there...

Jia-Qi Cao, Jinying Ma, De-Bin Zhou et al. · 0 citations
Open access 2026

Scale-Aware Fusion and Spatial-Frequency Collaborative Network for Remote Sensing Imagery Semantic Segmentation

Semantic segmentation of remote sensing images is crucial for serial earth observation tasks. However, significant scale variations in remote sensing scenes and insufficient exploitation of frequency-domain information often cause small-scale objects to be overwhelmed by large backgrounds under imbalanced multiscale fe...

Lu Wang, Chenxuan Lou, Jing Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.