EFCMF: A Multimodal Robustness Enhancement Framework for Fine-Grained Recognition

被引:0
|
作者
Zou, Rongping [1 ,2 ,3 ]
Zhu, Bin [1 ,2 ,3 ]
Chen, Yi [1 ,2 ,3 ]
Xie, Bo [1 ,2 ,3 ]
Shao, Bin [1 ,2 ,3 ]
机构
[1] Natl Univ Def Technol, Coll Elect Engn, Hefei 230037, Peoples R China
[2] State Key Lab Pulsed Power Laser Technol, Hefei 230037, Peoples R China
[3] Key Lab Infrared & Low Temp Plasma Anhui Prov, Hefei 230037, Peoples R China
来源
APPLIED SCIENCES-BASEL | 2023年 / 13卷 / 03期
基金
美国国家科学基金会;
关键词
fine-grained recognition; multimodal; modal missing; adversarial examples;
D O I
10.3390/app13031640
中图分类号
O6 [化学];
学科分类号
0703 ;
摘要
Fine-grained recognition has many applications in many fields and aims to identify targets from subcategories. This is a highly challenging task due to the minor differences between subcategories. Both modal missing and adversarial sample attacks are easily encountered in fine-grained recognition tasks based on multimodal data. These situations can easily lead to the model needing to be fixed. An Enhanced Framework for the Complementarity of Multimodal Features (EFCMF) is proposed in this study to solve this problem. The model's learning of multimodal data complementarity is enhanced by randomly deactivating modal features in the constructed multimodal fine-grained recognition model. The results show that the model gains the ability to handle modal missing without additional training of the model and can achieve 91.14% and 99.31% accuracy on Birds and Flowers datasets. The average accuracy of EFCMF on the two datasets is 52.85%, which is 27.13% higher than that of Bi-modal PMA when facing four adversarial example attacks, namely FGSM, BIM, PGD and C&W. In the face of missing modal cases, the average accuracy of EFCMF is 76.33% on both datasets respectively, which is 32.63% higher than that of Bi-modal PMA. Compared with existing methods, EFCMF is robust in the face of modal missing and adversarial example attacks in multimodal fine-grained recognition tasks. The source code is available at https://github.com/RPZ97/EFCMF (accessed on 8 January 2023).
引用
收藏
页数:15
相关论文
共 50 条
  • [11] Traffic Flow Model Informed Network for Fine-Grained Speed Estimation with Robustness Enhancement
    Li, Jin-Yu
    Tan, Hua-Chun
    Ding, Fan
    [J]. CICTP 2023: INNOVATION-EMPOWERED TECHNOLOGY FOR SUSTAINABLE, INTELLIGENT, DECARBONIZED, AND CONNECTED TRANSPORTATION, 2023, : 2907 - 2919
  • [12] Fine-grained Multimodal Entity Linking for Videos
    Zhao H.-Q.
    Wang X.-W.
    Li J.-L.
    Li Z.-X.
    Xiao Y.-H.
    [J]. Ruan Jian Xue Bao/Journal of Software, 2024, 35 (03): : 1140 - 1153
  • [13] Towards Fine-Grained Recognition: Joint Learning for Object Detection and Fine-Grained Classification
    Wang, Qiaosong
    Rasmussen, Christopher
    [J]. ADVANCES IN VISUAL COMPUTING, ISVC 2019, PT II, 2019, 11845 : 332 - 344
  • [14] SELECTIVE PARTS FOR FINE-GRAINED RECOGNITION
    Li, Dong
    Li, Yali
    Wang, Shengjin
    [J]. 2015 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING (ICIP), 2015, : 922 - 926
  • [15] FINE-GRAINED AND LAYERED OBJECT RECOGNITION
    Wu, Yang
    Zheng, Nanning
    Liu, Yuanliu
    Yuan, Zejian
    [J]. INTERNATIONAL JOURNAL OF PATTERN RECOGNITION AND ARTIFICIAL INTELLIGENCE, 2012, 26 (02)
  • [16] Deep LSAC for Fine-Grained Recognition
    Lin, Di
    Wang, Yi
    Liang, Lingyu
    Li, Ping
    Chen, C. L. Philip
    [J]. IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2022, 33 (01) : 200 - 214
  • [17] FgER: Fine-Grained Entity Recognition
    Abhishek
    [J]. THIRTY-SECOND AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE / THIRTIETH INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE CONFERENCE / EIGHTH AAAI SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2018, : 8008 - 8009
  • [18] A dataset for fine-grained seed recognition
    Yuan, Min
    Lv, Ningning
    Dong, Yongkang
    Hu, Xiaowen
    Lu, Fuxiang
    Zhan, Kun
    Shen, Jiacheng
    Wu, Xiaolin
    Zhu, Liye
    Xie, Yufei
    [J]. SCIENTIFIC DATA, 2024, 11 (01)
  • [19] Fine-Grained Differences-Similarities Enhancement Network for Multimodal Fake News Detection
    Wu, Xiaoyu
    Li, Shi
    Lai, Zhongyuan
    Song, Haifeng
    Hu, Chunfang
    [J]. INTERNATIONAL JOURNAL OF ADVANCED COMPUTER SCIENCE AND APPLICATIONS, 2023, 14 (10) : 1034 - 1042
  • [20] FUMMER: A fine-grained self-supervised momentum distillation framework for multimodal recommendation
    Wei, Yibiao
    Xu, Yang
    Zhu, Lei
    Ma, Jingwei
    Huang, Jiangping
    [J]. INFORMATION PROCESSING & MANAGEMENT, 2024, 61 (05)