Skip to content

MUMINS: Metadata-conditioned Uncertainty-aware Medical Image Next-state Synthesis

Sep 2026 · 1 citation · ⚡ 1 influential · 46 references
Computer Science

TL;DR

An efficient diffusion framework that jointly diffuses a baseline scan and its follow-up residual, summed to synthesize the follow-up scan, while concurrently predicting a spatial uncertainty map, in a single reverse diffusion process is proposed.

Abstract

Forecasting anatomical changes such as tumor growth and neurodegeneration is a challenging generative vision task. Morphological evolution is subtle relative to static anatomy, highly patient-specific, and inherently stochastic. Existing methods struggle with several issues: deterministic networks ignore biological stochasticity, while standard diffusion models require computationally prohibitive multi-pass sampling to quantify uncertainty. We propose MUMINS (Metadata-conditioned Uncertainty-aware Medical Image Next-state Synthesis), an efficient diffusion framework that jointly diffuses a baseline scan and its follow-up residual, summed to synthesize the follow-up scan, while concurrently predicting a spatial uncertainty map, in a single reverse diffusion process. Conditioned on the time interval and relevant metadata, it preserves fine-grained anatomy by dynamically re-injecting the baseline as a soft anchor at every denoising step, and a negative-log-likelihood head learns the uncertainty map to explicitly flag error-prone regions. Designed without organ-specific heuristics, the same architecture is reused across anatomies via separate, dataset-specific retraining. Extensive evaluations demonstrate that dataset-specific retraining of MUMINS matches or outperforms dedicated, domain-specific state-of-the-art methods on lung CT (PNG) and brain MRI (OASIS-3). Project page: https://github.com/aolivtous/MUMINS.

View source

Similar papers

Open access 2026

LISP-Net: Context-Aware and Lightweight Interactive Medical Segmentation

Precise medical image segmentation is essential to modern clinical workflows and biomedical research. However, current automated models often lack the flexibility, generalizability, and clinician control required to adapt to out-of-distribution data or novel classes without computationally expensive retraining. Further...

Paul Machauer, M. Reisert, Janis Keuper · 0 citations
Open access 2026

FSAUnet: Anisotropy-Aware Focal Modulation and Class-Aware MAC-Loss for 3-D Medical Image Segmentation

FSAUnet, a parameter-efficient three-dimensional segmentation framework built upon the self-configuring nnU-Net pipeline and featuring three innovations, obtains higher cross-validation mean Dice with fewer parameters under the evaluated settings, without establishing statistical significance, strict architecture-only...

Hong-Sen Yang, Ling-Fei Cheng, Yu-Xiang Liu · 0 citations
#diffusion models Open access Aug 2026

Robust unsupervised domain adaptation for medical image segmentation via frequency-conditioned graph diffusion

A novel structured latent UDA framework that performs domain alignment in a topology-aware representation space rather than directly modifying image appearance is proposed, highlighting the effectiveness of structured latent modeling and diffusion-based learning for robust domain-adaptive segmentation.

Usman Ahmad Usmani, Arunava Roy, Junzo Watada · 0 citations
Open access Sep 2026

Uncertainty-guided decoupled complementary network for binary semi-supervised medical image segmentation

Semi-supervised medical image segmentation methods have drawn wide attention as they reduce reliance on heavily annotated data. However, existing models suffer from confirmation bias with limited annotations, and structural or parameter coupling hinders self-correction, especially for medical images with ambiguous boun...

Dong-Sheng Wang, Xiao-Han Lang · 0 citations
Preprint Aug 2026

B-MIM: Biased Masked Image Modeling for Generalizable Segmentation of Fine-Grained Anatomical Structures

B-MIM is introduced, a modification of the iBOT objective that stochastically reduces global semantic alignment to prioritize local patch reconstruction and suggests that reducing global semantic pressure during pretraining enhances generalization to intricate anatomical structures.

S. González, Karen Sanchez, J. M. Saavedra et al. · 0 citations

Related blog posts

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.