Sep 2026· International Conference on Optics, Electronics, and Communication Engineering· Vol 14349, pp. 143492O - 143492O-7· 0 citations· 15 references
Engineering
TL;DR
These findings validate the effectiveness of combining CNNs and Transformer mechanisms in advancing automatic land use recognition and provide a promising pathway for scalable applications in large-scale remote sensing analysis.
Abstract
High-resolution remote sensing imagery provides rich spatial and semantic information for land use classification, which plays a crucial role in urban planning, resource management, and ecological monitoring. However, traditional convolutional neural network (CNN)-based approaches struggle to effectively capture long-range dependencies and complex contextual relationships inherent in such imagery. To address this limitation, we propose a Transformer-enhanced land use classification framework tailored for high-resolution remote sensing images. The method integrates local feature extraction via convolutional layers with global dependency modeling through Transformer attention mechanisms, thereby enabling both fine-grained texture recognition and contextual semantic understanding. Furthermore, a multi-scale feature fusion strategy and improved positional encoding are introduced to enhance representation robustness across varying resolutions. We evaluate the proposed model on four widely used benchmark datasets, including ISPRS Vaihingen, ISPRS Potsdam, UCM Land Use, and NWPU-RESISC45. Experimental results demonstrate that our approach achieves superior overall accuracy (OA), mean Intersection over Union (mIoU), and F1-scores compared to state-of-the-art baselines such as UNet, Deeplabv3+, and Swin Transformer. These findings validate the effectiveness of combining CNNs and Transformer mechanisms in advancing automatic land use recognition and provide a promising pathway for scalable applications in large-scale remote sensing analysis.
Remote sensing-based land cover classification and change detection are essential for ensuring environment, planning cities, and managing resources. Accurately extracting spatial and temporal information from high-resolution multi-temporal images remains challenging due to feature inconsistency, class imbalance, and...
Results indicate that the proposed HAA-UNet method effectively improves road continuity and boundary delineation in complex forest scenes and is integrated into the Forest Fire Risk Index (FFRI) assessment framework, demonstrating that accurate road data can improve the spatial characterization of fire risk and provide...
Hong-Rong Wang, Hao-Quan Chen, Fei-Fan Yang et al.· Sustainability· 0 citations
Reliable land-cover classification in arid oasis regions is imperative for pivotal decisions concerning oasis stability assessment, desertification control, and the optimal allocation of water resources. However, such regions are characterized by fragmented ground objects, inter-class homogeneity, intra-class heterogen...
Reliable mapping of intra-urban land-cover categories at fine spatial and thematic resolutions is fundamental to evidence-based spatial planning. Nevertheless, transferring classification models across territories remains difficult because urban morphology varies regionally. This challenge is pronounced in South Americ...
F. J. da Costa, R. Pescim, M. R. Urbano· Semina: Ciências Exatas e Te...· 0 citations
Road extraction from high-resolution remote sensing images is crucial for urban planning and geographic information systems (GIS). However, complex background interference, severe occlusions, and the inherent morphological complexity of roads often lead to discontinuities and insufficient accuracy in extraction results...
Jia-Jia Liu, Xuan Zhao, Wen-Xiang Dong et al.· Frontiers in Computing and I...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.