Skip to content

mmSimPrior: Learning Simulation Priors for Data-Efficient Real-World Generalizable Radar-Based Human Motion Reconstruction

Jul 2026 · arXiv.org · Vol abs/2607.22973 · 0 citations · 40 references
Computer Science

Abstract

Millimeter-wave (mmWave) radar enables privacy-preserving and illumination-robust human motion reconstruction, but training generalizable models typically requires costly paired radar-motion recordings. Simulation can scale such supervision, yet even physics-based simulators cannot fully reproduce real-world multipath, clutter, hardware-specific response statistics, or distance-dependent resolution degradation, leaving a sim-to-real gap. We present mmSimPrior, a simulation-pretrained framework that factorizes transferable knowledge into signal, motion, and radar-to-motion mapping priors. To learn transferable signal and motion priors, we pretrain a multimodal radar encoder with a physics-informed domain-randomization curriculum designed to mitigate the sim-to-real gap by approximating real-world propagation- and acquisition-level variations, while a joint-temporal tokenizer learns a discrete prior over plausible human motion. A dual-mode mapping module predicts either motion-code distributions for structurally constrained zero-shot reconstruction or continuous motion parameters for flexible adaptation from limited real data. We further construct a 4.2M-frame, 31K-sequence dataset suite and introduce a No-Overlap Setting that prevents any exact subject-environment-location-motion tuple from appearing in both the adaptation and test sets. Experiments on mmSimPrior-Real and RT-Pose demonstrate consistent gains: with only 24 paired real sequences, mmSimPrior-Reg reduces MPJPE by 24.7-39.0% over the strongest baseline across the three environments, while mmSimPrior-Cls reduces zero-shot MPJPE by 8.5% without fine-tuning.

View source

Similar papers

Preprint Jul 2026

HybridSim: A Physics-Learning Hybrid Digital Twin for mmWave Human Sensing

High-fidelity simulation of mmWave radar signals for dynamic human motion is valuable for developing radar-based human sensing models; yet collecting accurately labeled measurements for a specific deployment site remains expensive. We present HybridSim, a physics-learning hybrid simulator that synthesizes mmWave radar signals from dynamic human meshes under a fixed indoor room configuration, explicitly decoupling propagation into two components. To parameterize the human subject, we use a tri-plane representation to extract human features and a Graph Convolutional Network to stabilize optimization and mitigate gradient instability. The direct signal path is modeled via an inverse-rendering formulation with a microfacet BRDF to capture primary surface reflections. In parallel, the indirect path is approximated by combining 3D Gaussian Splatting with a virtual-receiver geometry to fit and reproduce site-specific multipath interference patterns, achieving substantially lower computational cost than explicit full ray tracing. Experiments in a fixed-room setting show improved agreement with a physically based reference and consistent gains on downstream radar-based human sensing tasks when using HybridSim for site-specific data augmentation.

Weitao Xiong, Tianyu Liu, Peng Li et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Physics-Unrolled Neural Operator for Wireless Field Modeling

This work proposes Physics-Unrolled Hybrid Neural Operator (PU-HNO), a three-stage cascade that predicts high-fidelity indoor radio maps from low-fidelity ray-tracing outputs and scene priors by progressively capturing reflection, diffraction, and scattering effects, rather than treating radio maps as generic images.

Rafid Umayer Murshed, S. ur Rahman, Mingyue Tang et al. · 0 citations
Preprint Aug 2026

You Only Flow Once: Calibrated and Real-Time Radar Pose Estimation with Multi-Hypothesis Normalizing Flows

This work proposes Multi-Hypothesis Normalizing Flow Pose Generator (MH-NFPG), which models pose distributions from radar point clouds using a conditional normalizing flow that transforms a Laplace base distribution into an expressive posterior, generated in parallel through a single forward pass.

J. Mueller, S. Hoefler, D. Zanca et al. · 0 citations
Preprint Aug 2026

CM-MAE: A Physics-Guided Cross-Modal Self-Supervised Learning Framework for Vision-Wireless Applications

CM-MAE is presented, a self-supervised vision--wireless pretraining framework for cross-scenario representation transfer that builds a target distribution from similarities between measured beam-power profiles, so nonidentical samples with similar directional responses are not forced apart as false negatives.

Yubo Zhang, Yi-Yao Liu · 0 citations
Open access 2026

Enhancing D-Band FMCW Radar Tracking for In-Air Writing Through Weakly Supervised Deep Association

The proposed TD-PDA generalizes to unseen users with an ultralow inference latency, successfully reconstructing legible trajectories even in the presence of strong multipath interference and achieves stability comparable to a well-tuned classical PDA filter via a purely data-driven design.

Salah Abouzaid, Leander Nothelle, Nils Pohl · 0 citations
Preprint Aug 2026

Simulation-Based Imaging: Learning Acoustic Inverse Problems from Simulated Data

We introduce Simulation-Based Imaging (SBI), a framework for non-destructive acoustic imaging in which machine learning models trained entirely on simulated data serve as real-time solvers for the acoustic inverse problem. A high-fidelity nodal Discontinuous Galerkin forward solver generates large training datasets by randomizing inclusion geometry within a unit-cube domain; a 2D convolutional neural network then learns a direct mapping from boundary pressure measurements to a 32 by 32 by 32 voxel reconstruction of the interior. The trained model reliably recovers inclusion position and size from 144 boundary sensors with no prior knowledge of inclusion count or geometry. Reconstruction error degrades by only 13% under 5% additive measurement noise, and just 17% of the sensor array (24 of 144 sensors) suffices for quality within 4% of full coverage. These results establish SBI as a viable proof-of-concept imaging device whose complexity resides in software rather than hardware, opening a path toward cheap, portable, deployable imaging systems.

L. Bodmer, E. Pitman · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.