Skip to content
Open access

Seed-net: structure-enhanced encoder-decoder network via dual-attention bridge for 2D medical image segmentation

Jul 2026 · Physica Scripta · Vol 101, pp. 306002 · 0 citations · 47 references
Physics

TL;DR

Results indicate that SEED-Net can effectively preserve fine structural details, reduce background interference, and produce accurate lesion boundaries, showing its potential for reliable MIS.

Abstract

Medical image segmentation (MIS) plays a crucial role in clinical diagnosis and disease prediction. However, existing encoder-decoder segmentation networks still face several structural limitations. First, conventional downsampling operations may discard high-frequency boundary details, resulting in blurred lesion edges. Second, shallow convolutional features are easily affected by local background noise and scale-variable lesion structures, while deep semantic representations often lack effective long-range dependency modeling, limiting global contextual understanding. Third, highly compressed bottleneck features often mix lesion semantics with background artifacts, which enlarges the semantic gap between the encoder and decoder. Finally, standard decoding operations may introduce upsampling artifacts and fail to accurately reconstruct irregular anatomical boundaries. To address these problems, we propose a structure-enhanced encoder-decoder network, termed SEED-Net, for 2D MIS. Specifically, Haar Wavelet Downsampling is introduced to preserve high-frequency structural information during feature compression, while Deep Mamba layers are deployed in deep semantic stages to capture long-range dependencies with linear computational complexity. In the encoder, the proposed encoder gradient multi-scale module introduces a purification-before-interaction strategy, where Grouped Dynamic Gating first recalibrates group-wise structural responses before local-global multi-scale feature interaction, thereby enhancing lesion-relevant structures while suppressing redundant background activations. At the bottleneck, the Large Kernel Dual-Attention Bridge recalibrates spatial and channel responses to isolate lesion-related semantics from background artifacts. In the decoder, the decoder gradient multi-scale Aggregation module improves boundary reconstruction and suppresses upsampling-induced artifacts through multi-receptive feature aggregation. Experimental results on five public datasets, including DSB2018, ISIC 2016, Kvasir-SEG, DRIVE,and LIDC-IDRI, demonstrate the effectiveness of SEED-Net. Specifically, on ISIC 2016, SEED-Net achieves an IoU of 85.54%, Dice of 91.61%, ACC of 95.64%, Spe of 96.25%, and Sen of 93.21%. These results indicate that SEED-Net can effectively preserve fine structural details, reduce background interference, and produce accurate lesion boundaries, showing its potential for reliable MIS.

Read PDF

Similar papers

Preprint Aug 2026

CiUNet: A Hybrid Swin-CNN UNet for Medical Image Segmentation

Medical image segmentation requires high accuracy and robustness, yet practical commercial deployment also demands privacy preservation and computational efficiency. In this context, the U-Net architecture, which can be inherently decoupled into independent encoder and decoder components, serves as a natural commercial...

Bin Dong, Jing-Hong Chen · 0 citations
Open access Aug 2026

TSPFusion: Tri-stream and prototype network for learning detail-semantic fusion in medical image segmentation

A tri-stream interaction paradigm replacing symmetric skip connections with directional fusion among semantic, spatial, and decoder-propagated streams at each decoding stage, and a Global Prototype Bank that captures dataset-level anatomical regularities via attention-based retrieval and gated EMA updates, providing pe...

Mohammed A. M. Elhassan, Qian-Fa Yuan, Zhizhong Xu et al. · 0 citations
Review Sep 2026

HFMD-Net: focal modulation and deformable convolutions for explainable pancreatic tumor segmentation with global–local boundary refinement

F focal-modulation-driven global context with deformable-convolution-based local adaptation produces boundary-refined pancreatic tumor segmentations on CT, which improves boundary fidelity and provides a transparent layer of review that may support treatment planning and longitudinal monitoring.

V. M. Firos, P. J. A. Alphonse, Ugo Fiore et al. · 0 citations
Open access Aug 2026

TransCat: a hybrid CNN-transformer network with KAN for medical image segmentation

TransCat, a hybrid CNN-Transformer architecture for medical image segmentation, is proposed and an extended deformable attention mechanism with attentive value identification is developed, to control the computational burden caused by the enlarged token set.

Jin Wang, Zheng-Hua Yang, Dong-Ming Zhou et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.