Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:VSFormer: Visual-Spatial Fusion Transformer for Correspondence Pruning

Jan 04, 2024

Tangfei Liao, Xiaoqin Zhang, Li Zhao, Tao Wang, Guobao Xiao

Figure 1 for VSFormer: Visual-Spatial Fusion Transformer for Correspondence Pruning

Figure 2 for VSFormer: Visual-Spatial Fusion Transformer for Correspondence Pruning

Figure 3 for VSFormer: Visual-Spatial Fusion Transformer for Correspondence Pruning

Figure 4 for VSFormer: Visual-Spatial Fusion Transformer for Correspondence Pruning

Share this with someone who'll enjoy it:

Abstract:Correspondence pruning aims to find correct matches (inliers) from an initial set of putative correspondences, which is a fundamental task for many applications. The process of finding is challenging, given the varying inlier ratios between scenes/image pairs due to significant visual differences. However, the performance of the existing methods is usually limited by the problem of lacking visual cues (\eg texture, illumination, structure) of scenes. In this paper, we propose a Visual-Spatial Fusion Transformer (VSFormer) to identify inliers and recover camera poses accurately. Firstly, we obtain highly abstract visual cues of a scene with the cross attention between local features of two-view images. Then, we model these visual cues and correspondences by a joint visual-spatial fusion module, simultaneously embedding visual cues into correspondences for pruning. Additionally, to mine the consistency of correspondences, we also design a novel module that combines the KNN-based graph and the transformer, effectively capturing both local and global contexts. Extensive experiments have demonstrated that the proposed VSFormer outperforms state-of-the-art methods on outdoor and indoor benchmarks. Our code is provided at the following repository: https://github.com/sugar-fly/VSFormer.

* Accepted by AAAI2024

View paper on

Share this with someone who'll enjoy it:

Title:VSFormer: Visual-Spatial Fusion Transformer for Correspondence Pruning

Paper and Code