Improving Small-Scale Human Action Recognition Performance Using a 3D Heatmap Volume-Reference-Cited by-同舟云学术

Improving Small-Scale Human Action Recognition Performance Using a 3D Heatmap Volume

Published:2023-07-13 Issue:14 Volume:23 Page:6364
ISSN:1424-8220
Container-title:Sensors
language:en
Short-container-title:Sensors

Author:

Yuan Lin¹,He Zhen¹,Wang Qiang¹^ORCID,Xu Leiyang¹^ORCID,Ma Xiang¹

Affiliation:

1. Department of Control Science and Engineering, Harbin Institute of Technology, Harbin 150001, China

Abstract

In recent years, skeleton-based human action recognition has garnered significant research attention, with proposed recognition or segmentation methods typically validated on large-scale coarse-grained action datasets. However, there remains a lack of research on the recognition of small-scale fine-grained human actions using deep learning methods, which have greater practical significance. To address this gap, we propose a novel approach based on heatmap-based pseudo videos and a unified, general model applicable to all modality datasets. Leveraging anthropometric kinematics as prior information, we extract common human motion features among datasets through an ad hoc pre-trained model. To overcome joint mismatch issues, we partition the human skeleton into five parts, a simple yet effective technique for information sharing. Our approach is evaluated on two datasets, including the public Nursing Activities and our self-built Tai Chi Action dataset. Results from linear evaluation protocol and fine-tuned evaluation demonstrate that our pre-trained model effectively captures common motion features among human actions and achieves steady and precise accuracy across all training settings, while mitigating network overfitting. Notably, our model outperforms state-of-the-art models in recognition accuracy when fusing joint and limb modality features along the channel dimension.

Funder

National Natural Science Foundation of China

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Biochemistry,Instrumentation,Atomic and Molecular Physics, and Optics,Analytical Chemistry

Link

https://www.mdpi.com/1424-8220/23/14/6364/pdf

Reference81 articles.

1. Patient activity recognition using radar sensors and machine learning;Bhavanasi;Neural Comput. Appl.,2022

2. A review of multimodal human activity recognition with special emphasis on classification, applications, challenges and future directions;Yadav;Knowl.-Based Syst.,2021

3. Soomro, K., Zamir, A.R., and Shah, M. (2012). UCF101: A dataset of 101 human actions classes from videos in the wild. arXiv.

4. Carreira, J., and Zisserman, A. (2017, January 21–26). Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset. Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.

5. Bertasius, G., Wang, H., and Torresani, L. (2021, January 18–24). Is Space-Time Attention All You Need for Video Understanding?. Proceedings of the 38th International Conference on Machine Learning, Online.

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. MAFormer: A cross-channel spatio-temporal feature aggregation method for human action recognition;AI Communications;2024-09-09

2. Basketball technique action recognition using 3D convolutional neural networks;Scientific Reports;2024-06-07

3. A Dense-Sparse Complementary Network for Human Action Recognition based on RGB and Skeleton Modalities;Expert Systems with Applications;2024-06