Feed-Forward Multi-view Multi-person Reconstruction with Contrastive Human-Aware 3D Representation
A new top-down paradigm is proposed that maintains a unified, instance-centric human-aware 3D space, enabling simultaneous camera calibration, cross-view association, and human reconstruction via cross-modal contrastive learning and allows correspondence reasoning, semantic aggregation, and instance discrimination to b...