On the Imaginary Wings: Text-Assisted Complex-Valued Fusion Network for Fine-Grained Visual Classification

被引:10
|
作者
Guan, Xiang [1 ]
Yang, Yang [1 ]
Li, Jingjing [1 ]
Zhu, Xiaofeng [1 ]
Song, Jingkuan [1 ]
Shen, Heng Tao [1 ,2 ]
机构
[1] Univ Elect Sci & Technol China, Ctr Future Media, Chengdu 611731, Peoples R China
[2] Peng Cheng Lab, Shenzhen 518066, Peoples R China
基金
中国国家自然科学基金;
关键词
Complex values; fine-grained visual classification (FGVC); graph convolutional networks (GCNs); multimodal;
D O I
10.1109/TNNLS.2021.3126046
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Fine-grained visual classification (FGVC) is challenging due to the interclass similarity and intraclass variation in datasets. In this work, we explore the great merit of complex values in introducing an imaginary part for modeling data uncertainty (e.g., different points on the complex plane can describe the same state) and graph convolutional networks (GCNs) in learning interdependently among classes to simultaneously tackle the above two major challenges. To the end, we propose a novel approach, termed text-assisted complex-valued fusion network (TA-CFN). Specifically, we expand each feature from 1-D real values to 2-D complex value by disassembling feature maps, thereby enabling the extension of traditional deep convolutional neural networks over the complex domain. Then, we fuse the real and imaginary parts of complex features through complex projection and modulus operation. Finally, we build an undirected graph over the object labels with the assistance of a text corpus, and a GCN is learned to map this graph into a set of classifiers. The benefits are in two folds: 1) complex features allow for a richer algebraic structure to better model the large variation within the same category and 2) leveraging the interclass dependencies brought by the GCN to capture key factors of the slight variation among different categories. We conduct extensive experiments to verify that our proposed model can achieve the state-of-the-art performance on two widely used FGVC datasets.
引用
收藏
页码:5112 / 5121
页数:10
相关论文
共 50 条
  • [21] Multiscale feature fusion and enhancement in a transformer for the fine-grained visual classification of tree species
    Dong, Yanqi
    Ma, Zhibin
    Zi, Jiali
    Xu, Fu
    Chen, Feixiang
    ECOLOGICAL INFORMATICS, 2025, 86
  • [22] Adaptive Feature Fusion Embedding Network for Few Shot Fine-Grained Image Classification
    Xie, Yaohua
    Zhang, Weichuan
    Ren, Jie
    Jing, Junfeng
    Computer Engineering and Applications, 2024, 59 (03) : 184 - 192
  • [23] Ship fine-grained classification network based on multi-scale feature fusion
    Chen, Lisu
    Wang, Qian
    Zhu, Enyan
    Feng, Daolun
    Wu, Huafeng
    Liu, Tao
    OCEAN ENGINEERING, 2025, 318
  • [24] Graph-in-graph discriminative feature enhancement network for fine-grained visual classification
    Wang, Yupeng
    Xu, Can
    Wang, Yongli
    Wang, Xiaoli
    Ding, Weiping
    APPLIED INTELLIGENCE, 2025, 55 (01)
  • [25] Multi-Granularity Feature Distillation Learning Network for Fine-Grained Visual Classification
    Cai, Yuhang
    Ke, Xiao
    2022 INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, COMPUTER VISION AND MACHINE LEARNING (ICICML), 2022, : 300 - 303
  • [26] A Feature Fusion Method Based on Multi-Classification Losses for Fine-Grained Visual Categorization
    Zhu, Mengmeng
    Wan, Shouhong
    Jin, Peiquan
    Tian, Qijun
    2021 IEEE INTERNATIONAL CONFERENCE ON BIG DATA (BIG DATA), 2021, : 6072 - 6074
  • [27] Significant feature suppression and cross-feature fusion networks for fine-grained visual classification
    Yang, Shengying
    Yang, Xinqi
    Wu, Jianfeng
    Feng, Boyang
    SCIENTIFIC REPORTS, 2024, 14 (01):
  • [28] Probability Fusion Decision Framework of Multiple Deep Neural Networks for Fine-Grained Visual Classification
    Zheng, Yang-Yang
    Kong, Jian-Lei
    Jin, Xue-Bo
    Wang, Xiao-Yi
    Su, Ting-Li
    Wang, Jian-Li
    IEEE ACCESS, 2019, 7 : 122740 - 122757
  • [29] Cross-layer progressive attention bilinear fusion method for fine-grained visual classification
    Wang, Chaoqing
    Qian, Yurong
    Gong, Weijun
    Cheng, Junjong
    Wang, Yongqiang
    Wang, Yuefei
    JOURNAL OF VISUAL COMMUNICATION AND IMAGE REPRESENTATION, 2022, 82
  • [30] I read, I saw, I tell: Texts Assisted Fine-Grained Visual Classification
    Li, Jingjing
    Zhu, Lei
    Huang, Zi
    Lu, Ke
    Zhao, Jidong
    PROCEEDINGS OF THE 2018 ACM MULTIMEDIA CONFERENCE (MM'18), 2018, : 663 - 671