AI-Generated Image Detection using a Cross-Attention Enhanced Dual-Stream Network

被引:0
|
作者
Xi, Ziyi [1 ]
Huang, Wenmin [1 ]
Wei, Kangkang [1 ]
Luo, Weiqi [1 ]
Zheng, Peijia [1 ]
机构
[1] Sun Yat Sen Univ, GuangDong Prov Key Lab Informat Secur Technol, Guangzhou, Peoples R China
关键词
D O I
10.1109/APSIPAASC58517.2023.10317126
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
With the rapid evolution of AI Generated Content (AIGC), forged images produced through this technology are inherently more deceptive and require less human intervention compared to traditional Computer-generated Graphics (CG). However, owing to the disparities between CG and AIGC, conventional CG detection methods tend to be inadequate in identifying AIGC-produced images. To address this issue, our research concentrates on the text-to-image generation process in AIGC. Initially, we first assemble two text-to-image databases utilizing two distinct AI systems, DALL center dot E2 and DreamStudio. Aiming to holistically capture the inherent anomalies produced by AIGC, we develope a robust dual-stream network comprised of a residual stream and a content stream. The former employs the Spatial Rich Model (SRM) to meticulously extract various texture information from images, while the latter seeks to capture additional forged traces in low frequency, thereby extracting complementary information that the residual stream may overlook. To enhance the information exchange between these two streams, we incorporate a cross multi-head attention mechanism. Numerous comparative experiments are performed on both databases, and the results show that our detection method consistently outperforms traditional CG detection techniques across a range of image resolutions. Moreover, our method exhibits superior performance through a series of robustness tests and cross-database experiments. When applied to widely recognized traditional CG benchmarks such as SPL2018 and DsTok, our approach significantly exceeds the capabilities of other existing methods in the field of CG detection.
引用
收藏
页码:1463 / 1470
页数:8
相关论文
共 50 条
  • [31] Multimodal Cross-Attention Graph Network for Desire Detection
    Gu, Ruitong
    Wang, Xin
    Yang, Qinghong
    ARTIFICIAL NEURAL NETWORKS AND MACHINE LEARNING, ICANN 2023, PT IV, 2023, 14257 : 512 - 523
  • [32] Dual-stream network with cross-layer attention and similarity constraint for micro-expression recognition
    Wang, Gang
    Huang, Shucheng
    MULTIMEDIA SYSTEMS, 2024, 30 (03)
  • [33] DCFNet: An Effective Dual-Branch Cross-Attention Fusion Network for Medical Image Segmentation
    Zhu, Chengzhang
    Zhang, Renmao
    Xiao, Yalong
    Zou, Beiji
    Chai, Xian
    Yang, Zhangzheng
    Hu, Rong
    Duan, Xuanchu
    CMES-COMPUTER MODELING IN ENGINEERING & SCIENCES, 2024, 140 (01): : 1103 - 1128
  • [34] CAFIN: cross-attention based face image repair network
    Li, Yaqian
    Li, Kairan
    Li, Haibin
    Zhang, Wenming
    MULTIMEDIA SYSTEMS, 2024, 30 (05)
  • [35] An Enhanced Cross-Attention Based Multimodal Model for Depression Detection
    Kou, Yifan
    Ge, Fangzhen
    Chen, Debao
    Shen, Longfeng
    Liu, Huaiyu
    Computational Intelligence, 2025, 41 (01)
  • [36] Dual-stream enhancement encoder and attention optimization decoder for image manipulation localization
    Zhu, Ye
    Zhao, Xiaoxiang
    Yu, Yang
    CHINESE JOURNAL OF LIQUID CRYSTALS AND DISPLAYS, 2024, 39 (08) : 1103 - 1115
  • [37] A Novel Transformer Network with a CNN-Enhanced Cross-Attention Mechanism for Hyperspectral Image Classification
    Wang, Xinyu
    Sun, Le
    Lu, Chuhan
    Li, Baozhu
    REMOTE SENSING, 2024, 16 (07)
  • [38] Quality Assessment of AI-Generated Image Based on Cross-modal Correlation
    Zhang, Yunhao
    Jia, Menglin
    Zhou, Wenbo
    Yang, Yang
    2024 3RD INTERNATIONAL CONFERENCE ON IMAGE PROCESSING AND MEDIA COMPUTING, ICIPMC 2024, 2024, : 378 - 384
  • [39] A WAVELET-BASED DUAL-STREAM NETWORK FOR UNDERWATER IMAGE ENHANCEMENT
    Ma, Ziyin
    Oh, Changjae
    2022 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP), 2022, : 2769 - 2773
  • [40] SwinTFNet: Dual-Stream Transformer With Cross Attention Fusion for Land Cover Classification
    Ren, Bo
    Liu, Bo
    Hou, Biao
    Wang, Zhao
    Yang, Chen
    Jiao, Licheng
    IEEE GEOSCIENCE AND REMOTE SENSING LETTERS, 2024, 21 : 1 - 5