Skip to content
Preprint

Ordered Diffusion for 3D Human Registration

Aug 2026 · 1 citation · 76 references
Computer Science

TL;DR

This work proposes ODin, which formulates registration as a 3D diffusion process that generates a point cloud aligned with the target geometry while preserving template semantics through consistent point ordering, and establishes a new state of the art in 3D human registration.

Abstract

3D human registration has historically been treated as a regression task, assuming a unique ground-truth alignment exists between the template and an input point cloud. In reality, acquisition noise, occlusions, and unknown soft tissue dynamics introduce inherent ambiguity into human scans. Regression-based methods consequently converge to an average prediction, often failing to represent a plausible geometry. In our work, we embrace such uncertainty by modeling the registration as a distribution of alignments. We propose ODin, which formulates registration as a 3D diffusion process that generates a point cloud aligned with the target geometry while preserving template semantics through consistent point ordering. To achieve this, ODin relies on global, local, and positional conditioning, guiding each point to its correct location. Our experiments demonstrate that such a generative formulation not only outperforms its regression-based baseline, but also establishes a new state of the art, surpassing highly engineered methods while reducing the registration time by two-thirds. Pre-trained models and code are available at https://riccardomarin.github.io/odin/.

View source

Similar papers

Preprint Aug 2026

DMM-Align: Closed-Loop Optimization for 2D-3D Registration with Dual-Role Diffusion

2D-3D registration remains brittle in challenging scenarios such as low overlap, occlusion, repetitive structures, and severe cross-modal ambiguity. A key reason is that existing methods improve representation learning, correspondence estimation, or pose computation in isolation, while the dominant failure mode is inhe...

Chong Wang, Jun-Jie Gao · 0 citations
Preprint Sep 2026

XPos3R: Cross-Modal Transformer for Intraoperative 2D/3D Registration

This work proposes XPos3R, a generalizable pose regression method that eliminates preoperative preparation, and introduces an asymmetric encoder-decoder architecture that improves cross-modal feature alignment while maintaining computational efficiency.

Shi-Yan Su, Ruyi Zha, Hong-Dong Li et al. · 0 citations
Preprint Aug 2026

The Right Prior for the Right Deformation: Rethinking Continuous Deformable Image Registration

Deformable image registration models implicitly encode deformation priors through their parametrization and optimization. In this work, we conduct a validation study on continuous registration methods to examine how these implicit priors affect performance across different registration tasks. Classic B-Spline transform...

Hengjie Liu, Chushu Shen, Dan Ruan et al. · 0 citations
Preprint Sep 2026

Evaluating Transformation Models for pCLE Mosaic Registration

Confocal Laser Endomicroscopy (CLE) provides real-time, cellular-resolution optical biopsy but has a narrow field of view, which image mosaicing can extend to provide anatomical context. Because of line-by-line acquisition, probe motion, and probe-tissue interaction, frame alignment generally requires a non-linear tran...

Ahmed Aboelela, J. Barcsay, Jana Friedhof et al. · 0 citations
Preprint Sep 2026

Geometry Without Coordinates: LiDAR Diffusion as a 3D Feature Bridge

Transferring the rich priors of large 2D foundation models to sparse 3D LiDAR remains challenging, as training native 3D foundation models at comparable scale is limited by data and annotation scarcity. We introduce a LiDAR-conditioned diffusion model trained on pseudo-labels from off-the-shelf 2D foundation models. Th...

Samed Doğan, Nico Leuze, Alfred Schöttl · 0 citations
Preprint Sep 2026

UBone3D: Physics-Rectified Conditional Flow Matching for Anatomical 3D Shape Completion from Ultrasound

Three-dimensional ultrasound (US) is a safe, radiation-free complementary modality to CT and X-rays for longitudinal monitoring, yet its segmentation-derived partial point clouds are extremely artifact-laden. Consequently, it is challenging to recover a clean and complete anatomical structure from such US point clouds....

Wei-Ying Chen, Yuchong Gao, Si-Yuan Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.