Multi-layered semantic representation network for multi-label image classification

被引:0
|
作者
Xiwen Qu
Hao Che
Jun Huang
Linchuan Xu
Xiao Zheng
机构
[1] Anhui University of Technology,School of Computer Science and Technology
[2] Hefei Comprehensive National Science Center,Institute of Artificial Intelligence
[3] Australian National University,Department of Computing
[4] The Hong Kong Polytechnic University,undefined
关键词
Multi-label image classification; Convolutional neural network; Label embeddings; Multi-layered attention;
D O I
暂无
中图分类号
学科分类号
摘要
Multi-label image classification is a fundamental and practical task, which aims to assign multiple possible labels to an image. In recent years, many deep convolutional neural network (CNN) based approaches have been proposed which model label correlations to discover semantics of labels and learn semantic representations of images. This paper advances this research direction by improving both the modeling of label correlations and the learning of semantic representations. On the one hand, besides the local semantics of each label, we propose to further explore global semantics shared by multiple labels. On the other hand, existing approaches mainly learn the semantic representations at the last convolutional layer of a CNN. But it has been noted that the image representations of different layers of CNN capture different levels or scales of features and have different discriminative abilities. We thus propose to learn semantic representations at multiple convolutional layers. To this end, this paper designs a Multi-layered Semantic Representation Network (MSRN) which discovers both local and global semantics of labels through modeling label correlations and utilizes the label semantics to guide the semantic representations learning at multiple layers through an attention mechanism. Extensive experiments on five benchmark datasets including VOC2007, VOC2012, MS-COCO, NUS-WIDE, and Apparel show a competitive performance of the proposed MSRN against state-of-the-art models.
引用
下载
收藏
页码:3427 / 3435
页数:8
相关论文
共 50 条
  • [31] Video Representation Fusion Network For Multi-Label Movie Genre Classification
    Bi, Tianyu
    Jarnikov, Dmitri
    Lukkien, Johan
    2020 25TH INTERNATIONAL CONFERENCE ON PATTERN RECOGNITION (ICPR), 2021, : 9386 - 9391
  • [32] Multi-layered image representation: Application to image compression
    Meyer, FG
    Averbuch, AZ
    Stromberg, JO
    Coifman, RR
    1998 INTERNATIONAL CONFERENCE ON IMAGE PROCESSING - PROCEEDINGS, VOL 2, 1998, : 292 - 296
  • [33] Multi-Label Fake News Detection using Multi-layered Supervised Learning
    Rasool, Tayyaba
    Butt, Wasi Haider
    Shaukat, Arslan
    Akram, M. Usman
    PROCEEDINGS OF 2019 11TH INTERNATIONAL CONFERENCE ON COMPUTER AND AUTOMATION ENGINEERING (ICCAE 2019), 2019, : 73 - 77
  • [34] Multi-label classification of traditional national costume pattern image semantic understanding
    Zhao H.-Y.
    Zhou W.
    Hou X.-G.
    Qi G.-L.
    Guangxue Jingmi Gongcheng/Optics and Precision Engineering, 2020, 28 (03): : 695 - 703
  • [35] Multiple Semantic Embedding with Graph Convolutional Networks for Multi-Label Image Classification
    Zhou, Tong
    Feng, Songhe
    PATTERN RECOGNITION AND COMPUTER VISION, PRCV 2021, PT II, 2021, 13020 : 449 - 461
  • [36] MULTI-LABEL TEXT CLASSIFICATION WITH A ROBUST LABEL DEPENDENT REPRESENTATION
    Alfaro, Rodrigo
    Allende, Hector
    2011 INTERNATIONAL CONFERENCE ON INSTRUMENTATION, MEASUREMENT, CIRCUITS AND SYSTEMS (ICIMCS 2011), VOL 3: COMPUTER-AIDED DESIGN, MANUFACTURING AND MANAGEMENT, 2011, : 211 - 214
  • [37] Latent Semantic Aware Multi-View Multi-Label Classification
    Zhang, Changqing
    Yu, Ziwei
    Hu, Qinghua
    Zhu, Pengfei
    Liu, Xinwang
    Wang, Xiaobo
    THIRTY-SECOND AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE / THIRTIETH INNOVATIVE APPLICATIONS OF ARTIFICIAL INTELLIGENCE CONFERENCE / EIGHTH AAAI SYMPOSIUM ON EDUCATIONAL ADVANCES IN ARTIFICIAL INTELLIGENCE, 2018, : 4414 - 4421
  • [38] Attention-Augmented Memory Network for Image Multi-Label Classification
    Zhou, Wei
    Hou, Yanke
    Chen, Dihu
    Hu, Haifeng
    Su, Tao
    ACM TRANSACTIONS ON MULTIMEDIA COMPUTING COMMUNICATIONS AND APPLICATIONS, 2023, 19 (03)
  • [39] General Multi-label Image Classification with Transformers
    Lanchantin, Jack
    Wang, Tianlu
    Ordonez, Vicente
    Qi, Yanjun
    2021 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR 2021, 2021, : 16473 - 16483
  • [40] Double Attention for Multi-Label Image Classification
    Zhao, Haiying
    Zhou, Wei
    Hou, Xiaogang
    Zhu, Hui
    IEEE ACCESS, 2020, 8 : 225539 - 225550