KFSENet: A Key Frame-Based Skeleton Feature Estimation and Action Recognition Network for Improved Robot Vision with Face and Emotion Recognition-Reference-Cited by-同舟云学术

KFSENet: A Key Frame-Based Skeleton Feature Estimation and Action Recognition Network for Improved Robot Vision with Face and Emotion Recognition

Published:2022-05-27 Issue:11 Volume:12 Page:5455
ISSN:2076-3417
Container-title:Applied Sciences
language:en
Short-container-title:Applied Sciences

Author:

Le Dinh-Son,Phan Hai-Hong^ORCID,Hung Ha Huy^ORCID,Tran Van-An,Nguyen The-Hung,Nguyen Dinh-Quan^ORCID

Abstract

In this paper, we propose an integrated approach to robot vision: a key frame-based skeleton feature estimation and action recognition network (KFSENet) that incorporates action recognition with face and emotion recognition to enable social robots to engage in more personal interactions. Instead of extracting the human skeleton features from the entire video, we propose a key frame-based approach for their extraction using pose estimation models. We select the key frames using the gradient of a proposed total motion metric that is computed using dense optical flow. We use the extracted human skeleton features from the selected key frames to train a deep neural network (i.e., the double-feature double-motion network (DDNet)) for action recognition. The proposed KFSENet utilizes a simpler model to learn and differentiate between the different action classes, is computationally simpler and yields better action recognition performance when compared with existing methods. The use of key frames allows the proposed method to eliminate unnecessary and redundant information, which improves its classification accuracy and decreases its computational cost. The proposed method is tested on both publicly available standard benchmark datasets and self-collected datasets. The performance of the proposed method is compared to existing state-of-the-art methods. Our results indicate that the proposed method yields better performance compared with existing methods. Moreover, our proposed framework integrates face and emotion recognition to enable social robots to engage in more personal interaction with humans.

Publisher

MDPI AG

Subject

Fluid Flow and Transfer Processes,Computer Science Applications,Process Chemistry and Technology,General Engineering,Instrumentation,General Materials Science

Link

https://www.mdpi.com/2076-3417/12/11/5455/pdf

Reference50 articles.

1. Reinforcement Learning Approaches in Social Robotics

2. A review of recent research in social robotics

3. Preparing for a robot future? Social professions, social robotics and the challenges ahead;Share;Ir. J. Appl. Soc. Stud.,2018

4. A Survey of Vision-Based Human Action Evaluation Methods

Cited by 6 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. MDJ: A multi-scale difference joint keyframe extraction algorithm for infrared surveillance video action recognition;Digital Signal Processing;2024-05

2. Various frameworks for integrating image and video streams for spatiotemporal information learning employing 2D–3D residual networks for human action recognition;Discover Applied Sciences;2024-03-18

3. Enhancing Image Clarity: Feature Selection with Trickster Coyote Optimization in Noisy/Blurry Images;Salud, Ciencia y Tecnología;2024-01-01

4. Analyzing audiovisual data for understanding user's emotion in human−computer interaction environment;Data Technologies and Applications;2023-11-01

5. New Trends in Emotion Recognition Using Image Analysis by Neural Networks, a Systematic Review;Sensors;2023-08-10