Guided CNN for generalized zero-shot and open-set recognition using visual and semantic prototypes

被引:32
|
作者
Geng, Chuanxing [1 ]
Tao, Lue [1 ]
Chen, Songcan [1 ]
机构
[1] Nanjing Univ Aeronaut & Astronaut, Coll Comp Sci & Technol, MIIT Key Lab Pattern Anal & Machine Intelligence, Nanjing 211106, Peoples R China
关键词
Convolutional prototype learning; Generalized zero-shot Learning; Open set recognition;
D O I
10.1016/j.patcog.2020.107263
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
In the process of exploring the world, the curiosity constantly drives humans to cognize new things. Supposing you are a zoologist, for a presented animal image, you can recognize it immediately if you know its class. Otherwise, you would more likely attempt to cognize it by exploiting the side-information (e.g., semantic information, etc.) you have accumulated. Inspired by this, this paper decomposes the generalized zero-shot learning (G-ZSL) task into an open set recognition (OSR) task and a zero-shot learning (ZSL) task, where OSR recognizes seen classes (if we have seen (or known) them) and rejects unseen classes (if we have never seen (or known) them before), while ZSL identifies the unseen classes rejected by the former. Simultaneously, without violating OSR's assumptions (only known class knowledge is available in training), we also first attempt to explore a new generalized open set recognition (G-OSR) by introducing the accumulated side-information from known classes to OSR. For G-ZSL, such a decomposition effectively solves the class overfitting problem with easily misclassifying unseen classes as seen classes. The problem is ubiquitous in most existing G-ZSL methods. On the other hand, for G-OSR, introducing such semantic information of known classes not only improves the recognition performance but also endows OSR with the cognitive ability of unknown classes. Specifically, a visual and semantic prototypes-jointly guided convolutional neural network (VSG-CNN) is proposed to fulfill these two tasks (G-ZSL and G-OSR) in a unified end-to-end learning framework. Extensive experiments on benchmark datasets demonstrate the advantages of our learning framework. (C) 2020 Elsevier Ltd. All rights reserved.
引用
收藏
页数:10
相关论文
共 50 条
  • [1] Counterfactual Zero-Shot and Open-Set Visual Recognition
    Yue, Zhongqi
    Wang, Tan
    Sun, Qianru
    Hua, Xian-Sheng
    Zhang, Hanwang
    2021 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR 2021, 2021, : 15399 - 15409
  • [2] Joint Feature Generation and Open-set Prototype Learning for generalized zero-shot open-set classification
    Li, Xiao
    Fang, Min
    Zhai, Zhibo
    PATTERN RECOGNITION, 2024, 147
  • [3] Indirect visual–semantic alignment for generalized zero-shot recognition
    Yan-He Chen
    Mei-Chen Yeh
    Multimedia Systems, 2024, 30
  • [4] Exemplar-Based, Semantic Guided Zero-Shot Visual Recognition
    Zhang, Chunjie
    Liang, Chao
    Zhao, Yao
    IEEE TRANSACTIONS ON IMAGE PROCESSING, 2022, 31 : 3056 - 3065
  • [5] Indirect visual-semantic alignment for generalized zero-shot recognition
    Chen, Yan-He
    Yeh, Mei-Chen
    MULTIMEDIA SYSTEMS, 2024, 30 (02)
  • [6] CoHOZ: Contrastive Multimodal Prompt Tuning for Hierarchical Open-set Zero-shot Recognition
    Liao, Ning
    Liu, Yifeng
    Li, Xiaobo
    Lei, Chenyi
    Wang, Guoxin
    Hua, Xian-Sheng
    Yan, Junchi
    PROCEEDINGS OF THE 30TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA, MM 2022, 2022, : 3262 - 3271
  • [7] Vocabulary-Informed Zero-Shot and Open-Set Learning
    Fu, Yanwei
    Wang, Xiaomei
    Dong, Hanze
    Jiang, Yu-Gang
    Wang, Meng
    Xue, Xiangyang
    Sigal, Leonid
    IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2020, 42 (12) : 3136 - 3152
  • [8] GENERALIZED ZERO-SHOT RECOGNITION THROUGH IMAGE-GUIDED SEMANTIC CLASSIFICATION
    Li, Fang
    Yeh, Mei-Chen
    2021 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING (ICIP), 2021, : 2483 - 2487
  • [9] Exploring Zero-Shot Emotion Recognition in Speech Using Semantic-Embedding Prototypes
    Xu, Xinzhou
    Deng, Jun
    Cummins, Nicholas
    Zhang, Zixing
    Zhao, Li
    Schuller, Bjoern W.
    IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 24 : 2752 - 2765
  • [10] Generalized Zero-Shot Recognition based on Visually Semantic Embedding
    Zhu, Pengkai
    Wang, Hanxiao
    Saligrama, Venkatesh
    2019 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2019), 2019, : 2990 - 2998