GeoFF3D reconstructs 2,000 images in approximately five minutes, demonstrating scalable and robust large-scale UAV reconstruction, which combines a coordinate-anchored model with a spatial large-scale reconstruction framework (SLRF).
Abstract
Existing feed-forward 3D reconstruction methods typically process a bounded number of images and recover cameras and geometry in local or internally normalized frames. Extending them to large-scale UAV mapping requires scalable multi-chunk processing and reliable aggregation, while full Sim(3) alignment can become unstable for near collinear trajectories. We present GeoFF3D, which combines a coordinate-anchored model with a spatial large-scale reconstruction framework (SLRF). The model uses georeferenced camera translations and optional geometric priors to predict camera poses and dense point maps directly in a gravity-aligned Z-up metric frame. SLRF partitions images into spatially overlapping chunks, propagates shared-view priors, and aggregates local reconstructions hierarchically, while remaining applicable to different bounded-view models. Across nine aerial mapping blocks, GeoFF3D achieves the best average reconstruction quality, improving F@5 from 0.829 for Pi3X + SLRF to 0.877. On long UAVScenes sequences, it reaches 0.848, compared with 0.687 for Pi3X + SLRF and 0.451 for the strongest evaluated SLAM/streaming baseline. GeoFF3D reconstructs 2,000 images in approximately five minutes, demonstrating scalable and robust large-scale UAV reconstruction.The code is available at https://github.com/yanxian-ll/GeoFF3D.
An iterative hybrid discrete-continuous viewpoint planning method for targeted UAV photogrammetry from a proxy reconstruction that improves both reconstruction accuracy and completeness compared with prior UAV path-planning methods.
Alana Grech, Daniel Pisani, Andrew Grima et al.· 0 citations
Ground image localization with respect to satellite imagery is a key enabler for metrically-accurate, geo-localized 3D scene reconstruction from unconstrained image collections. Existing cross-view localization methods have strict requirements such as panoramic imagery or known initial locations, limiting their applica...
A. Daruna, Ben Southall, Niluthpol Chowdhury Mithun et al.· 0 citations
While feed-forward 3D reconstruction (3R) offers efficient end-to-end modeling, its application in large-scale UAV mapping is hindered by the prohibitive memory cost of Transformer attention. Current scalable streaming 3R methods assume temporally and spatially continuous inputs, rendering them ineffective for the weak...
Zhe Shen, Liyuan Lou, Yifei Yu et al.· 0 citations
Unmanned aerial vehicles (UAVs) are increasingly utilized across military, civilian, and agricultural sectors, necessitating accurate and efficient 3D target localization. Traditional 2D detectors lack depth perception, while stereo vision and LiDAR have range-dependent and hardware limitations, respectively. To addres...
Yan-Xin Sun, Ming-Ming Ma, Lan-Yu Sun et al.· Electronics· 0 citations
GeoWeaver, a unified framework comprising a Geometric Prior Model (GPM) and Test-Time Adaptation (TTA) and a robust CDF-style objective jointly optimizes weighted 2D reprojection and 3D consistency residuals, is presented.
Tinghao Jiang, Sheng Tang, Shengzhe Wei et al.· 0 citations
Feed-forward 3D foundation models reconstruct perspective scenes in one pass. Satellite photogrammetry needs a different product, one that domain adaptation alone does not deliver: dense surface height in an absolute geodetic frame under non-central rational polynomial cameras (RPCs). Perspective-pretrained features ar...
Zhe Dong, Wan-Qin Wu, Yuzhe Sun et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.