跳到主要导航 跳到搜索 跳到主要内容

TransVLAD: Multi-Scale Attention-Based Global Descriptors for Visual Geo-Localization

  • Yifan Xu*
  • , Pourya Shamsolmoali
  • , Eric Granger
  • , Claire Nicodeme
  • , Laurent Gardes
  • , Jie Yang
  • *此作品的通讯作者
  • Shanghai Jiao Tong University
  • École de technologie supérieure
  • SNCF

科研成果: 书/报告/会议事项章节会议稿件同行评审

摘要

Visual geo-localization remains a challenging task due to variations in the appearance and perspective among captured images. This paper introduces an efficient TransVLAD module, which aggregates attention-based feature maps into a discriminative and compact global descriptor. Unlike existing methods that generate feature maps using only convolutional neural networks (CNNs), we propose a sparse transformer to encode global dependencies and compute attention-based feature maps, which effectively reduces visual ambiguities that occurs in large-scale geo-localization problems. A positional embedding mechanism is used to learn the corresponding geometric configurations between query and gallery images. A grouped VLAD layer is also introduced to reduce the number of parameters, and thus construct an efficient module. Finally, rather than only learning from the global descriptors on entire images, we propose a self-supervised learning method to further encode more information from multi-scale patches between the query and positive gallery images. Extensive experiments on three challenging large-scale datasets indicate that our model outperforms state-of-the-art models, and has lower computational complexity. The code is available at: https://github.com/wacv-23/TVLAD.

源语言英语
主期刊名Proceedings - 2023 IEEE Winter Conference on Applications of Computer Vision, WACV 2023
出版商Institute of Electrical and Electronics Engineers Inc.
2839-2848
页数10
ISBN(电子版)9781665493468
DOI
出版状态已出版 - 2023
活动23rd IEEE/CVF Winter Conference on Applications of Computer Vision, WACV 2023 - Waikoloa, 美国
期限: 3 1月 20237 1月 2023

出版系列

姓名Proceedings - 2023 IEEE Winter Conference on Applications of Computer Vision, WACV 2023

会议

会议23rd IEEE/CVF Winter Conference on Applications of Computer Vision, WACV 2023
国家/地区美国
Waikoloa
时期3/01/237/01/23

指纹

探究 'TransVLAD: Multi-Scale Attention-Based Global Descriptors for Visual Geo-Localization' 的科研主题。它们共同构成独一无二的指纹。

引用此