Two-stage Discriminative Re-ranking for Large-scale Landmark Retrieval

被引:4
|
作者
Yokoo, Shuhei [1 ]
Ozaki, Kohei [2 ,4 ]
Simo-Serra, Edgar [3 ]
Iizuka, Satoshi [1 ]
机构
[1] Univ Tsukuba, Tsukuba, Ibaraki, Japan
[2] Preferred Networks Inc, Tokyo, Japan
[3] Waseda Univ, Tokyo, Japan
[4] Recruit Technol Co Ltd, Tokyo, Japan
关键词
QUERY EXPANSION; IMAGE; DESCRIPTORS; FEATURES; GEOMETRY; MODEL;
D O I
10.1109/CVPRW50498.2020.00514
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
We propose an efficient pipeline for large-scale landmark image retrieval that addresses the diversity of the dataset through two-stage discriminative re-ranking. Our approach is based on embedding the images in a feature-space using a convolutional neural network trained with a cosine softmax loss. Due to the variance of the images, which include extreme viewpoint changes such as having to retrieve images of the exterior of a landmark from images of the interior, this is very challenging for approaches based exclusively on visual similarity. Our proposed re-ranking approach improves the results in two steps: in the sort-step, k-nearest neighbor search with soft-voting to sort the retrieved results based on their label similarity to the query images, and in the insert-step, we add additional samples from the dataset that were not retrieved by image-similarity. This approach allows overcoming the low visual diversity in retrieved images. In-depth experimental results show that the proposed approach significantly outperforms existing approaches on the challenging Google Landmarks Datasets. Using our methods, we achieved 1st place in the Google Landmark Retrieval 2019 challenge on Kaggle. Our code is publicly available here: https://github.com/lyakaap/Landmark2019-1st-and-3rd-Place-Solution
引用
收藏
页码:4363 / 4370
页数:8
相关论文
共 50 条
  • [41] A two-stage optimization strategy for large-scale oil field development
    Yusuf Nasir
    Oleg Volkov
    Louis J. Durlofsky
    Optimization and Engineering, 2022, 23 : 361 - 395
  • [42] Combining re-ranking and rank aggregation methods for image retrieval
    Guimaraes Pedronette, Daniel Carlos
    Torres, Ricardo da S.
    MULTIMEDIA TOOLS AND APPLICATIONS, 2016, 75 (15) : 9121 - 9144
  • [43] Two-Stage Nonnegative Sparse Representation for Large-Scale Face Recognition
    He, Ran
    Zheng, Wei-Shi
    Hu, Bao-Gang
    Kong, Xiang-Wei
    IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2013, 24 (01) : 35 - 46
  • [44] A two-stage design for multiple testing in large-scale association studies
    Wen, Shu-Hui
    Tzeng, Jung-Ying
    Kao, Jau-Tsuen
    Hsiao, Chuhsing Kate
    JOURNAL OF HUMAN GENETICS, 2006, 51 (06) : 523 - 532
  • [45] FAST GEOMETRIC RE-RANKING FOR IMAGE-BASED RETRIEVAL
    Tsai, Sam S.
    Chen, David
    Takacs, Gabriel
    Chandrasekhar, Vijay
    Vedantham, Ramakrishna
    Grzeszczuk, Radek
    Girod, Bernd
    2010 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, 2010, : 1029 - 1032
  • [46] Visual Re-Ranking for Multi-Aspect Information Retrieval
    Klouche, Khalil
    Ruotsalo, Tuukka
    Micallef, Luana
    Andolina, Salvatore
    Jacucci, Giulio
    CHIIR'17: PROCEEDINGS OF THE 2017 CONFERENCE HUMAN INFORMATION INTERACTION AND RETRIEVAL, 2017, : 57 - 66
  • [47] Graph Convolution Based Efficient Re-Ranking for Visual Retrieval
    Zhang, Yuqi
    Qian, Qi
    Wang, Hongsong
    Liu, Chong
    Chen, Weihua
    Wang, Fan
    IEEE TRANSACTIONS ON MULTIMEDIA, 2024, 26 : 1089 - 1101
  • [48] Efficient Re-ranking in Vocabulary Tree based Image Retrieval
    Wang, Xiaoyu
    Yang, Ming
    Yu, Kai
    2011 CONFERENCE RECORD OF THE FORTY-FIFTH ASILOMAR CONFERENCE ON SIGNALS, SYSTEMS & COMPUTERS (ASILOMAR), 2011, : 855 - 859
  • [49] Adaptive Query Re-ranking Based on ImageGraph for Image Retrieval
    Fan, Haonan
    Hu, Hai-Miao
    Wang, Rong
    Zhang, Yugui
    2018 IEEE INTERNATIONAL CONFERENCE ON BIG DATA (BIG DATA), 2018, : 4593 - 4599
  • [50] Combining re-ranking and rank aggregation methods for image retrieval
    Daniel Carlos Guimarães Pedronette
    Ricardo da S. Torres
    Multimedia Tools and Applications, 2016, 75 : 9121 - 9144