Decomposed Soft Prompt Guided Fusion Enhancing for Compositional Zero-Shot Learning

被引:12
|
作者
Lu, Xiaocheng [1 ]
Guo, Song [1 ,2 ]
Liu, Ziming [1 ]
Guo, Jingcai [1 ,2 ]
机构
[1] Hong Kong Polytech Univ, Dept Comp, Hong Kong, Peoples R China
[2] Hong Kong Polytech Univ, Shenzhen Res Inst, Hong Kong, Peoples R China
基金
中国国家自然科学基金;
关键词
D O I
10.1109/CVPR52729.2023.02256
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Compositional Zero-Shot Learning (CZSL) aims to recognize novel concepts formed by known states and objects during training. Existing methods either learn the combined state-object representation, challenging the generalization of unseen compositions, or design two classifiers to identify state and object separately from image features, ignoring the intrinsic relationship between them. To jointly eliminate the above issues and construct a more robust CZSL system, we propose a novel framework termed Decomposed Fusion with Soft Prompt (DFSP)1, by involving vision-language models (VLMs) for unseen composition recognition. Specifically, DFSP constructs a vector combination of learnable soft prompts with state and object to establish the joint representation of them. In addition, a cross-modal decomposed fusion module is designed between the language and image branches, which decomposes state and object among language features instead of image features. Notably, being fused with the decomposed features, the image features can be more expressive for learning the relationship with states and objects, respectively, to improve the response of unseen compositions in the pair space, hence narrowing the domain gap between seen and unseen sets. Experimental results on three challenging benchmarks demonstrate that our approach significantly outperforms other state-of-the-art methods by large margins.
引用
收藏
页码:23560 / 23569
页数:10
相关论文
共 50 条
  • [21] ProZe: Explainable and Prompt-Guided Zero-Shot Text Classification
    Harrando, Ismail
    Reboud, Alison
    Schleider, Thomas
    Ehrhart, Thibault
    Troncy, Raphael
    IEEE INTERNET COMPUTING, 2022, 26 (06) : 69 - 77
  • [22] Generating Variable Explanations via Zero-shot Prompt Learning
    Wang, Chong
    Lou, Yiling
    Liu, Junwei
    Peng, Xin
    2023 38TH IEEE/ACM INTERNATIONAL CONFERENCE ON AUTOMATED SOFTWARE ENGINEERING, ASE, 2023, : 748 - 760
  • [23] KG-SP: Knowledge Guided Simple Primitives for OpenWorld Compositional Zero-Shot Learning
    Karthik, Shyamgopal
    Mancini, Massimiliano
    Akata, Zeynep
    2022 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2022, : 9326 - 9335
  • [24] Learning Graph Embeddings for Open World Compositional Zero-Shot Learning
    Mancini, Massimiliano
    Naeem, Muhammad Ferjad
    Xian, Yongqin
    Akata, Zeynep
    IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2024, 46 (03) : 1545 - 1560
  • [25] Hierarchical Visual Primitive Experts for Compositional Zero-Shot Learning
    Kim, Hanjae
    Lee, Jiyoung
    Park, Seongheon
    Sohn, Kwanghoon
    2023 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION, ICCV, 2023, : 5652 - 5662
  • [26] Siamese Contrastive Embedding Network for Compositional Zero-Shot Learning
    Li, Xiangyu
    Yang, Xu
    Wei, Kun
    Deng, Cheng
    Yang, Muli
    2022 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2022, : 9316 - 9325
  • [27] Swap-Reconstruction Autoencoder for Compositional Zero-Shot Learning
    Guo, Ting
    Liang, Jiye
    Xie, Guo-Sen
    2023 IEEE INTERNATIONAL CONFERENCE ON MULTIMEDIA AND EXPO, ICME, 2023, : 438 - 443
  • [28] Enhancing Classification in Zero-Shot Learning with the Aid of Perceptron
    Zengin, Hilal
    Ismailoglu, Firat
    2022 30TH SIGNAL PROCESSING AND COMMUNICATIONS APPLICATIONS CONFERENCE, SIU, 2022,
  • [29] FFusion: Feature Fusion Transformer for Zero-Shot Learning
    Tao, Wenjin
    Xie, Jiahao
    An, Zhinan
    Meng, Xianjia
    ELECTRONICS, 2025, 14 (05):
  • [30] Zero-Shot Rumor Detection with Propagation Structure via Prompt Learning
    Lin, Hongzhan
    Yi, Pengyao
    Ma, Jing
    Jiang, Haiyun
    Luo, Ziyang
    Shi, Shuming
    Liu, Ruifang
    THIRTY-SEVENTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE, VOL 37 NO 4, 2023, : 5213 - 5221