Skip to content
Preprint

Surfsvr: 2D Surface Priors as 3D Geometric Regularizers for Sparse Voxel Reconstruction

Aug 2026 · 0 citations · 58 references
Computer Science

Abstract

Sparse voxel reconstruction offers an efficient representation for high-fidelity 3D modeling, yet its geometry is commonly optimized from local photometric evidence and discrete visibility statistics. This often leads to fragmented surfaces, excessive subdivision, and floating artifacts, particularly in weakly textured or sparsely observed regions. We introduce SurfSVR, a novel sparse voxel reconstruction paradigm that treats 2D surface priors as explicit 3D geometric regularizers. Instead of directly lifting noisy pixel-wise depth predictions, SurfSVR first organizes each image into coherent surface regions by jointly reasoning over appearance, monocular depth, normals and cross-view geometry. Each region is then represented by an adaptively selected planar or quadratic surface model based on fitting reliability and geometric complexity, while cross-model agreement distinguishes reliable geometry from ambiguous predictions. These structured 2D priors are lifted into 3D and integrated throughout the reconstruction pipeline. They guide surface-adaptive voxel subdivision, provide region-level depth and normal supervision during optimization, enhance geometrically reliable sparse-observed surfaces in voxel pruning, and suppress off-surface floaters during post-refinement training. This unified design converts semantic and geometric coherence in image space into persistent structural constraints in 3D. Extensive experiments on 3 public benchmarks demonstrate that SurfSVR consistently improves sparse voxel reconstruction across scenes with substantially different visibility and geometry characteristics, achieving state-of-the-art reconstruction quality. Codes and models will be released soon.

View source

Similar papers

Preprint Aug 2026

Point-Based 3D Reconstruction from Sparse Views under Known Illumination

These results show that, in the controlled direct illumination setting, compact beta surfels combined with transport-based optimization can recover surfaces without relying on the tens to hundreds of thousands of primitives used by the evaluated baselines.

M. K. Gjerde, J. B. Haurum, J. Frisvad et al. · 0 citations
Preprint Sep 2026

AnyGS2Mesh: Feed-Forward Mesh Reconstruction from 3D Gaussian Splatting with Arbitrary-Resolution Views

AnyGS2Mesh is presented, the first feed-forward framework for directly reconstructing 3D meshes from 3D Gaussian Splatting representations with support for arbitrary input image resolutions, and demonstrates the potential of combining Gaussian representations and feed-forward Transformer architectures for scalable 3D g...

Yuxuan Song, Fan Gao, Yi-Bo Zhao et al. · 0 citations
Preprint Sep 2026

Superquadric Primitive Decomposition of 3D point clouds via Geometric-Aware Inlier Refinement

The decomposition of 3D point clouds into interpretable geometric primitives remains a longstanding challenge in Computer Vision and Computer Graphics. Among the available representations, superquadrics offer a compact and expressive model capable of capturing a wide range of shapes. However, their estimation is inhere...

Alessandro Rinaldi, Edoardo Tedesco, Andrea Ferraris et al. · 0 citations
Open access Aug 2026

High-Fidelity Gaussian Splatting from MVS Clouds: An Iterative Spatial Decomposition Framework

This work proposes an Iterative Spatial Decomposition framework that bridges dense geometric priors from Multi-View Stereo (MVS) with Gaussian Splatting and introduces Hierarchical Geometric Prior Sampling (HGPS), which substantially reduce redundancy in MVS point clouds while preserving critical details, thereby provi...

Zong-Hua Yu, Jun-Huai Li, Huai-Jun Wang et al. · 0 citations
Open access Sep 2026

SurfelFlow: Surface-Aware 2D Gaussian Streaming for Monocular Dynamic 4D Reconstruction

Recovering time-varying 3D scenes from monocular dynamic videos is challenging because each frame provides only a single view of a changing scene. Fast, large-magnitude motion further weakens geometric stability and image fidelity under monocular observations, especially when depth, motion, and visibility must be infer...

Qiao-Lian Xue, Yu Zhong, Ming-Qiang Xu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.