Integrating Human Parsing and Pose Network for Human Action Recognition

被引:0
|
作者
Ding, Runwei [1 ]
Wen, Yuhang [2 ]
Liu, Jinfu [2 ]
Dai, Nan [3 ]
Meng, Fanyang [4 ]
Liu, Mengyuan [1 ]
机构
[1] Peking Univ, Shenzhen Grad Sch, Shenzhen, Peoples R China
[2] Sun Yat Sen Univ, Shenzhen, Peoples R China
[3] Changchun Univ Sci & Technol, Changchun, Peoples R China
[4] Peng Cheng Lab, Shenzhen, Peoples R China
来源
基金
中国国家自然科学基金;
关键词
Action recognition; Human parsing; Human skeletons;
D O I
10.1007/978-981-99-8850-1_15
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Human skeletons and RGB sequences are both widelyadopted input modalities for human action recognition. However, skeletons lack appearance features and color data suffer large amount of irrelevant depiction. To address this, we introduce human parsing feature map as a novel modality, since it can selectively retain spatiotemporal features of the body parts, while filtering out noises regarding outfits, backgrounds, etc. We propose an Integrating Human Parsing and Pose Network (IPP-Net) for action recognition, which is the first to leverage both skeletons and human parsing feature maps in dual-branch approach. The human pose branch feeds compact skeletal representations of different modalities in graph convolutional network to model pose features. In human parsing branch, multi-frame body-part parsing features are extracted with human detector and parser, which is later learnt using a convolutional backbone. A late ensemble of two branches is adopted to get final predictions, considering both robust keypoints and rich semantic body-part features. Extensive experiments on NTU RGB+D and NTU RGB+D 120 benchmarks consistently verify the effectiveness of the proposed IPP-Net, which outperforms the existing action recognition methods. Our code is publicly available at https://github.com/liujf69/IPPNet-Parsing.
引用
收藏
页码:182 / 194
页数:13
相关论文
共 50 条
  • [31] Human action recognition in videos with articulated pose information by deep networks
    Farrajota, M.
    Rodrigues, Joao M. F.
    du Buf, J. M. H.
    PATTERN ANALYSIS AND APPLICATIONS, 2019, 22 (04) : 1307 - 1318
  • [32] Human action recognition using Pose-based discriminant embedding
    Saghafi, Behrouz
    Rajan, Deepu
    SIGNAL PROCESSING-IMAGE COMMUNICATION, 2012, 27 (01) : 96 - 111
  • [33] On-Line Human Action Recognition by Combining Joint Tracking and Key Pose Recognition
    Weng, E-Jui
    Fu, Li-Chen
    2012 IEEE/RSJ INTERNATIONAL CONFERENCE ON INTELLIGENT ROBOTS AND SYSTEMS (IROS), 2012, : 4112 - 4117
  • [34] Discriminative Hierarchical Part-based Models for Human Parsing and Action Recognition
    Wang, Yang
    Duan Tran
    Liao, Zicheng
    Forsyth, David
    JOURNAL OF MACHINE LEARNING RESEARCH, 2012, 13 : 3075 - 3102
  • [35] Integrating Joint and Surface for Human Action Recognition in Indoor Environments
    Li, Qingyang
    Zhou, Yu
    Ming, Anlong
    2014 INTERNATIONAL CONFERENCE ON SECURITY, PATTERN ANALYSIS, AND CYBERNETICS (SPAC), 2014, : 100 - 104
  • [36] SUNNet: A novel framework for simultaneous human parsing and pose estimation
    Xu, Yanyu
    Piao, Zhixin
    Zhang, Ziheng
    Liu, Wen
    Gao, Shenghua
    NEUROCOMPUTING, 2021, 444 : 349 - 355
  • [37] Human Action Invarianceness for Human Action Recognition
    Sjarif, Nilam Nur Amir
    Shamsuddin, Siti Mariyam
    2015 9TH INTERNATIONAL CONFERENCE ON SOFTWARE, KNOWLEDGE, INFORMATION MANAGEMENT AND APPLICATIONS (SKIMA), 2015,
  • [38] Neural Architecture Search for Joint Human Parsing and Pose Estimation
    Zeng, Dan
    Huang, Yuhang
    Bao, Qian
    Zhang, Junjie
    Su, Chi
    Liu, Wu
    2021 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV 2021), 2021, : 11365 - 11374
  • [39] An efficient human action recognition framework with pose-based spatiotemporal features
    Agahian, Saeid
    Negin, Farhood
    Kose, Cemal
    ENGINEERING SCIENCE AND TECHNOLOGY-AN INTERNATIONAL JOURNAL-JESTECH, 2020, 23 (01): : 196 - 203
  • [40] Automatic Key Pose Selection for 3D Human Action Recognition
    Gong, Wenjuan
    Bagdanov, Andrew D.
    Xavier Roca, F.
    Gonzalez, Jordi
    ARTICULATED MOTION AND DEFORMABLE OBJECTS, PROCEEDINGS, 2010, 6169 : 290 - 299