Skip to content

VSLG-net: visual-spatial latent graph network for two-view correspondence learning

Aug 2026 · Signal, Image and Video Processing · Vol 20 · 0 citations · 50 references

TL;DR

The Visual-Spatial Latent Graph Network (VSLG-Net), a parameter-compact transformer-based framework with dual-branch for local and global context perception in attention mechanisms, is proposed, which achieves competitive performance on the outdoor YFCC100M benchmark and remains competitive on the indoor SUN3D benchmark, compared with state-of-the-art methods.

View source

Similar papers

Conference Sep 2026

BAG-Net: Bidirectional Receptive-Field Graph Network for Two-View Correspondence Pruning

Learning reliable two-view correspondences is essential for geometric computer vision applications. Existing graph-based pruning methods typically aggregate information unidirectionally, capturing only whether a correspondence receives neighborhood support while ignoring its structural contribution during aggregation....

Le-Yi Wang, Hao Chen, Chang-Cai Yang · 0 citations
#artificial intelligence Preprint Sep 2026

VGGT-Diff: Visual Geometry Meets Diffusion for Sparse-View Novel View Synthesis

This work presents VGGT-Diff, a geometry-routed multi-view diffusion model for sparse-view novel view synthesis, and introduces robust geometry conditioning, combining training-time regularization with inference-time guidance for improved robustness.

Kang-Jie Chen, Xiang-Yu Li, Dong-Bin Zhang et al. · 0 citations
Aug 2026

Graph-driven contextual synergy network for robust 3D object detection.

A Graph-driven Contextual Synergy Network (GCS3D), which is designed to systematically enhance point representations across both semantic and geometric dimensions, and introduces a Graph-guided Geometric Consistency Interaction module for contextual correlation modeling.

Miao-Hui Zhang, Cheng-Yi Zhang, Lin-Xian Zhu et al. · 0 citations
Open access Aug 2026

Semantic Topological Multi-Scale Part Network for Fine-Grained Visual Classification

A Prior-Guided Part Aggregator (PGA) is designed, which leverages the foreground prior provided by foundation models to guide discriminative part discovery, enhancing target region responses while suppressing background interference, and a Topology-Informed Semantic Graph Convolutional Network (TIS-GCN) is designed to...

Xue-Rong Liu, Min Zhi, Yan-Jun Yin et al. · 0 citations
Sep 2026

VLMSNet: View Large to Measure Shape

Recent advances in single image super-resolution (SISR) have leveraged convolutional neural networks (CNNs) and vision transformers to model pixel-level statistics, often relying on increasingly complex architectures to capture spatial correlations. However, these approaches generally overlook a fundamental distinction...

Qi-Bin Zhang, Li-Cheng Liu, Ting-Yun Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.