Integrating Human Parsing and Pose Network for Human Action Recognition

被引:0
|
作者
Ding, Runwei [1 ]
Wen, Yuhang [2 ]
Liu, Jinfu [2 ]
Dai, Nan [3 ]
Meng, Fanyang [4 ]
Liu, Mengyuan [1 ]
机构
[1] Peking Univ, Shenzhen Grad Sch, Shenzhen, Peoples R China
[2] Sun Yat Sen Univ, Shenzhen, Peoples R China
[3] Changchun Univ Sci & Technol, Changchun, Peoples R China
[4] Peng Cheng Lab, Shenzhen, Peoples R China
来源
基金
中国国家自然科学基金;
关键词
Action recognition; Human parsing; Human skeletons;
D O I
10.1007/978-981-99-8850-1_15
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Human skeletons and RGB sequences are both widelyadopted input modalities for human action recognition. However, skeletons lack appearance features and color data suffer large amount of irrelevant depiction. To address this, we introduce human parsing feature map as a novel modality, since it can selectively retain spatiotemporal features of the body parts, while filtering out noises regarding outfits, backgrounds, etc. We propose an Integrating Human Parsing and Pose Network (IPP-Net) for action recognition, which is the first to leverage both skeletons and human parsing feature maps in dual-branch approach. The human pose branch feeds compact skeletal representations of different modalities in graph convolutional network to model pose features. In human parsing branch, multi-frame body-part parsing features are extracted with human detector and parser, which is later learnt using a convolutional backbone. A late ensemble of two branches is adopted to get final predictions, considering both robust keypoints and rich semantic body-part features. Extensive experiments on NTU RGB+D and NTU RGB+D 120 benchmarks consistently verify the effectiveness of the proposed IPP-Net, which outperforms the existing action recognition methods. Our code is publicly available at https://github.com/liujf69/IPPNet-Parsing.
引用
收藏
页码:182 / 194
页数:13
相关论文
共 50 条
  • [1] Human Action Recognition Based on Integrating Body Pose, Part Shape, and Motion
    El-Ghaish, Hany
    Hussien, Mohamed E.
    Shoukry, Amin
    Onai, Rikio
    [J]. IEEE ACCESS, 2018, 6 : 49040 - 49055
  • [2] Explore human parsing modality for action recognition
    Liu, Jinfu
    Ding, Runwei
    Wen, Yuhang
    Dai, Nan
    Meng, Fanyang
    Zhang, Fang-Lue
    Zhao, Shen
    Liu, Mengyuan
    [J]. CAAI TRANSACTIONS ON INTELLIGENCE TECHNOLOGY, 2024,
  • [3] A Pose-Aware Global Representation Network for Human Parsing
    Zhou, Yanghong
    Mok, P. Y.
    [J]. IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2023, 33 (04) : 1710 - 1724
  • [4] Discriminative Pose Analysis for Human Action Recognition
    Zhao, Xiaofeng
    Huang, Yao
    Yang, Jianyu
    Liu, Chunping
    [J]. 2020 IEEE 6TH WORLD FORUM ON INTERNET OF THINGS (WF-IOT), 2020,
  • [5] Human action recognition by leaning pose dictionary
    Cai, Jiaxin
    Feng, Guocan
    Tang, Xin
    Luo, Zhihong
    [J]. Cai, Jiaxin, 1600, Chinese Optical Society (34):
  • [6] Learning pose dictionary for human action recognition
    Cai, Jia-xin
    Tang, Xin
    Feng, Guo-can
    [J]. 2014 22ND INTERNATIONAL CONFERENCE ON PATTERN RECOGNITION (ICPR), 2014, : 381 - 386
  • [7] Lightweight human pose estimation network and angle-based action recognition
    Kang, Guang-Yu
    Lu, Zi-Qian
    Lu, Zhe-Ming
    [J]. Journal of Network Intelligence, 2020, 5 (04): : 240 - 250
  • [8] Pose graph parsing network for human-object interaction detection
    Su, Zhan
    Wang, Yuting
    Xie, Qing
    Yu, Ruiyun
    [J]. NEUROCOMPUTING, 2022, 476 : 53 - 62
  • [9] Human pose estimation and action recognition for fitness movements☆
    Fu, Huichen
    Gao, Junwei
    Liu, Huabo
    [J]. COMPUTERS & GRAPHICS-UK, 2023, 116 : 418 - 426
  • [10] Pose Guided Dynamic Image Network for Human Action Recognition in Person Centric Videos
    Chaudhary, Sachin
    Dudhane, Akshay
    Patil, Prashant
    Murala, Subrahmanyam
    [J]. 2019 16TH IEEE INTERNATIONAL CONFERENCE ON ADVANCED VIDEO AND SIGNAL BASED SURVEILLANCE (AVSS), 2019,