Two-stage Discriminative Re-ranking for Large-scale Landmark Retrieval

被引:4
|
作者
Yokoo, Shuhei [1 ]
Ozaki, Kohei [2 ,4 ]
Simo-Serra, Edgar [3 ]
Iizuka, Satoshi [1 ]
机构
[1] Univ Tsukuba, Tsukuba, Ibaraki, Japan
[2] Preferred Networks Inc, Tokyo, Japan
[3] Waseda Univ, Tokyo, Japan
[4] Recruit Technol Co Ltd, Tokyo, Japan
关键词
QUERY EXPANSION; IMAGE; DESCRIPTORS; FEATURES; GEOMETRY; MODEL;
D O I
10.1109/CVPRW50498.2020.00514
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
We propose an efficient pipeline for large-scale landmark image retrieval that addresses the diversity of the dataset through two-stage discriminative re-ranking. Our approach is based on embedding the images in a feature-space using a convolutional neural network trained with a cosine softmax loss. Due to the variance of the images, which include extreme viewpoint changes such as having to retrieve images of the exterior of a landmark from images of the interior, this is very challenging for approaches based exclusively on visual similarity. Our proposed re-ranking approach improves the results in two steps: in the sort-step, k-nearest neighbor search with soft-voting to sort the retrieved results based on their label similarity to the query images, and in the insert-step, we add additional samples from the dataset that were not retrieved by image-similarity. This approach allows overcoming the low visual diversity in retrieved images. In-depth experimental results show that the proposed approach significantly outperforms existing approaches on the challenging Google Landmarks Datasets. Using our methods, we achieved 1st place in the Google Landmark Retrieval 2019 challenge on Kaggle. Our code is publicly available here: https://github.com/lyakaap/Landmark2019-1st-and-3rd-Place-Solution
引用
收藏
页码:4363 / 4370
页数:8
相关论文
共 50 条
  • [21] Discriminative re-ranking for automatic speech recognition by leveraging invariant structures
    Suzuki, Masayuki
    Kurata, Gakuto
    Nishimura, Masafumi
    Minematsu, Nobuaki
    SPEECH COMMUNICATION, 2015, 72 : 208 - 217
  • [22] Efficient Large-Scale Video Retrieval via Discriminative Signatures
    Hao, Pengyi
    Kamata, Sei-ichiro
    IEICE TRANSACTIONS ON INFORMATION AND SYSTEMS, 2013, E96D (08) : 1800 - 1810
  • [23] RE-RANKING WITH TWO-LEVEL HASHING
    Feng, Shao-Yong
    Tian, Xing
    Ng, Wing W. Y.
    2014 INTERNATIONAL CONFERENCE ON WAVELET ANALYSIS AND PATTERN RECOGNITION (ICWAPR), 2014, : 98 - 103
  • [24] Accurate Image Retrieval with Unsupervised 2-Stage k-NN Re-Ranking
    Li, Dawei
    Chuah, Mooi Choo
    INTERNATIONAL JOURNAL OF MULTIMEDIA DATA ENGINEERING & MANAGEMENT, 2016, 7 (01): : 41 - 59
  • [25] Two-Stage Sparse Representation for Robust Recognition on Large-Scale Database
    He, Ran
    Hu, BaoGang
    Zheng, Wei-Shi
    Guo, YanQing
    PROCEEDINGS OF THE TWENTY-FOURTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE (AAAI-10), 2010, : 475 - 480
  • [26] A two-stage design for multiple testing in large-scale association studies
    Shu-Hui Wen
    Jung-Ying Tzeng
    Jau-Tsuen Kao
    Chuhsing Kate Hsiao
    Journal of Human Genetics, 2006, 51 : 523 - 532
  • [27] Testing successive regression approximations by large-scale two-stage problems
    Deak, Istvan
    ANNALS OF OPERATIONS RESEARCH, 2011, 186 (01) : 83 - 99
  • [28] Two-Stage Precoding Method for the Finitely Large-Scale Antenna Systems
    Joonwoo Shin
    Wireless Personal Communications, 2015, 84 : 2549 - 2559
  • [29] A Two-Stage Vehicle Routing Model for Large-Scale Bioterrorism Emergencies
    Shen, Zhihong
    Dessouky, Maged M.
    Ordonez, Fernando
    NETWORKS, 2009, 54 (04) : 255 - 269
  • [30] Optimization of natural frequencies of large-scale two-stage raft system
    Lv Zhiqiang
    He Lin
    Shuai Changgeng
    13TH INTERNATIONAL CONFERENCE ON MOTION AND VIBRATION CONTROL (MOVIC 2016) AND THE 12TH INTERNATIONAL CONFERENCE ON RECENT ADVANCES IN STRUCTURAL DYNAMICS (RASD 2016), 2016, 744