On the Imaginary Wings: Text-Assisted Complex-Valued Fusion Network for Fine-Grained Visual Classification

被引:10
|
作者
Guan, Xiang [1 ]
Yang, Yang [1 ]
Li, Jingjing [1 ]
Zhu, Xiaofeng [1 ]
Song, Jingkuan [1 ]
Shen, Heng Tao [1 ,2 ]
机构
[1] Univ Elect Sci & Technol China, Ctr Future Media, Chengdu 611731, Peoples R China
[2] Peng Cheng Lab, Shenzhen 518066, Peoples R China
基金
中国国家自然科学基金;
关键词
Complex values; fine-grained visual classification (FGVC); graph convolutional networks (GCNs); multimodal;
D O I
10.1109/TNNLS.2021.3126046
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Fine-grained visual classification (FGVC) is challenging due to the interclass similarity and intraclass variation in datasets. In this work, we explore the great merit of complex values in introducing an imaginary part for modeling data uncertainty (e.g., different points on the complex plane can describe the same state) and graph convolutional networks (GCNs) in learning interdependently among classes to simultaneously tackle the above two major challenges. To the end, we propose a novel approach, termed text-assisted complex-valued fusion network (TA-CFN). Specifically, we expand each feature from 1-D real values to 2-D complex value by disassembling feature maps, thereby enabling the extension of traditional deep convolutional neural networks over the complex domain. Then, we fuse the real and imaginary parts of complex features through complex projection and modulus operation. Finally, we build an undirected graph over the object labels with the assistance of a text corpus, and a GCN is learned to map this graph into a set of classifiers. The benefits are in two folds: 1) complex features allow for a richer algebraic structure to better model the large variation within the same category and 2) leveraging the interclass dependencies brought by the GCN to capture key factors of the slight variation among different categories. We conduct extensive experiments to verify that our proposed model can achieve the state-of-the-art performance on two widely used FGVC datasets.
引用
收藏
页码:5112 / 5121
页数:10
相关论文
共 50 条
  • [11] Complex Scenario-Oriented Fine-Grained Visual Classification Platform
    Chang, Dongliang
    Chen, Junhan
    Wang, Xinran
    Du, Ruoyi
    Yu, Wenqing
    Liu, Yufan
    Tong, Yujun
    Liang, Kongming
    Song, Yi-Zhe
    Ma, Zhanyu
    2022 IEEE 24TH INTERNATIONAL WORKSHOP ON MULTIMEDIA SIGNAL PROCESSING (MMSP), 2022,
  • [12] Progressive Co-Attention Network for Fine-Grained Visual Classification
    Zhang, Tian
    Chang, Dongliang
    Ma, Zhanyu
    Guo, Jun
    2021 INTERNATIONAL CONFERENCE ON VISUAL COMMUNICATIONS AND IMAGE PROCESSING (VCIP), 2021,
  • [13] Progressive Erasing Network with consistency loss for fine-grained visual classification
    Peng, Jin
    Wang, Yongxiong
    Zhou, Zeping
    JOURNAL OF VISUAL COMMUNICATION AND IMAGE REPRESENTATION, 2022, 87
  • [14] Multi-directional guidance network for fine-grained visual classification
    Yang, Shengying
    Jin, Yao
    Lei, Jingsheng
    Zhang, Shuping
    VISUAL COMPUTER, 2024, 40 (11): : 8113 - 8124
  • [15] Multi-modal hierarchical fusion network for fine-grained paper classification
    Tan Yue
    Yong Li
    Jiedong Qin
    Zonghai Hu
    Multimedia Tools and Applications, 2024, 83 : 31527 - 31543
  • [16] Complemental Attention Multi-Feature Fusion Network for Fine-Grained Classification
    Miao, Zhuang
    Zhao, Xun
    Wang, Jiabao
    Li, Yang
    Li, Hang
    IEEE SIGNAL PROCESSING LETTERS, 2021, 28 : 1983 - 1987
  • [17] Multi-modal hierarchical fusion network for fine-grained paper classification
    Yue, Tan
    Li, Yong
    Qin, Jiedong
    Hu, Zonghai
    MULTIMEDIA TOOLS AND APPLICATIONS, 2024, 83 (11) : 31527 - 31543
  • [18] Feature fusion network based on few-shot fine-grained classification
    Yang, Yajie
    Feng, Yuxuan
    Zhu, Li
    Fu, Haitao
    Pan, Xin
    Jin, Chenlei
    FRONTIERS IN NEUROROBOTICS, 2023, 17
  • [19] Multi-level navigation network: advancing fine-grained visual classification
    Liang, Hong
    Li, Xian
    Shao, Mingwen
    Zhang, Qian
    JOURNAL OF SUPERCOMPUTING, 2025, 81 (02):
  • [20] Fine-Grained Visual Classification Based on Sparse Bilinear Convolutional Neural Network
    Ma L.
    Wang Y.
    Moshi Shibie yu Rengong Zhineng/Pattern Recognition and Artificial Intelligence, 2019, 32 (04): : 336 - 344