Sep 2026· IEEE Transactions on Neural Networks and Learning Systems· Vol PP, pp. 1-14· 0 citations
Medicine
TL;DR
SRWKV is proposed, a shape-guided RWKV (SGR) model for parameter-efficient MedISeg that introduces an SGR block that uses a shape prior predicted from the deepest encoder feature to guide token traversal during decoding, reducing foreground-background interleaving and improving structural coherence during sequence formation.
Abstract
Existing medical image segmentation (MedISeg) models predominantly rely on convolutional neural networks (CNNs) and Transformer architectures. However, the limited receptive fields of CNNs and the quadratic computational cost of Transformers hinder their scalability and efficiency. Recently, receptance-weighted key-value (RWKV) has emerged as a promising linear-complexity alternative for global context modeling. In this article, we propose SRWKV, a shape-guided RWKV (SGR) model for parameter-efficient MedISeg. SRWKV introduces an SGR block that uses a shape prior predicted from the deepest encoder feature to guide token traversal during decoding, reducing foreground-background interleaving and improving structural coherence during sequence formation. In addition, we develop a deformable adaptive shift (DA-Shift) module that dynamically adjusts token interactions according to local context, enabling flexible receptive field adaptation for diverse anatomical structures. Extensive experiments across six MedISeg tasks on 11 datasets demonstrate that SRWKV achieves strong segmentation performance with a compact parameter footprint. Our code is available at https://github.com/ukeLin/SRWKV.
TvaraNet is pro-posed, an extremely lightweight segmentation network designed to preserve boundary fidelity under strict efficiency constraints and achieves competitive or superior boundary-aware performance compared to heavier architectures.
Sridhatta Jayaram Aithal, Vandana Bharti· Proceedings of the Thirty-Fi...· 0 citations
Multi-dimensional lightweight modules can reconcile segmentation quality with strict computational budgets when adapting video-centric foundation models to 2D clinical data.
Xue-Jia Yuan, Zong-Jian Yang, Yu Guo et al.· IEEE transactions on bio-med...· 0 citations
Medical image segmentation requires high accuracy and robustness, yet practical commercial deployment also demands privacy preservation and computational efficiency. In this context, the U-Net architecture, which can be inherently decoupled into independent encoder and decoder components, serves as a natural commercial...
Accurate kidney and renal-tumor segmentation is challenging because lesion size, location, morphology, and boundary contrast vary substantially across abdominal CT scans. Most existing methods rely on a single form of local evidence and struggle to recover the boundaries of small lesions while maintaining global anat...
Si-Yuan Liang, Cheng-Chuan Xu, Chao Lu et al.· Frontiers in Artificial Inte...· 0 citations
Introduction State Space Models (SSMs) have demonstrated strong potential for 3D brain tumor segmentation owing to their linear computational complexity. However, conventional Mamba-based models are often limited by spectral bias, which favors low-frequency information while overlooking high-frequency boundary details,...
Xiang-Ning Hou, Jun Yao, Qiao-Chu Li et al.· Frontiers in Oncology· 0 citations
F focal-modulation-driven global context with deformable-convolution-based local adaptation produces boundary-refined pancreatic tumor segmentations on CT, which improves boundary fidelity and provides a transparent layer of review that may support treatment planning and longitudinal monitoring.
V. M. Firos, P. J. A. Alphonse, Ugo Fiore et al.· Machine Vision and Applicati...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.