Video hand gestures recognition using depth camera and lightweight CNN

González León, David; Gröli, Jade; Reddy Yeduri, Sreenivasa; Rossier, Daniel; Mosqueron, Romuald; Pandey, Om Jee; Reddy Cenkeramaddi, Linga

doi:10.1109/JSEN.2022.3181518

González León, David; Gröli, Jade; Reddy Yeduri, Sreenivasa; Rossier, Daniel; Mosqueron, Romuald; Pandey, Om Jee; Reddy Cenkeramaddi, Linga

2022

Télécharger

Formats

Format
BibTeX
MARCXML
TextMARC
MARC
DublinCore
EndNote
NLM
RefWorks
RIS

Résumé

Hand gestures are a well-known and intuitive method of human-computer interaction. The majority of the research has concentrated on hand gesture recognition from the RGB images, however, little work has been done on recognition from videos. In addition, RGB cameras are not robust in varying lighting conditions. Motivated by this, we present the video based hand gestures recognition using the depth camera and a light weight convolutional neural network (CNN) model. We constructed a dataset and then used a light weight CNN model to detect and classify hand movements efficiently. We also examined the classification accuracy with a limited number of frames in a video gesture. We compare the depth camera’s video gesture recognition performance to that of the RGB camera. We evaluate the proposed model’s performance on edge computing devices and compare to benchmark models in terms of accuracy and inference time. The proposed model results in an accuracy of 99.48% on the RGB version of test dataset and 99.18% on the depth version of test dataset. Finally, we compare the accuracy of the proposed light weight CNN model with the state-of-the hand gesture classification models.

Détails

Titre Video hand gestures recognition using depth camera and lightweight CNN

Auteur(s)/ trice(s) González León, David (School of Engineering and Management Vaud, HES-SO, University of Applied Sciences and Arts Western Switzerland)
Gröli, Jade (School of Engineering and Management Vaud, HES-SO, University of Applied Sciences and Arts Western Switzerland)
Reddy Yeduri, Sreenivasa (University of Agder, Grimstad, Norway)
Rossier, Daniel (School of Engineering and Management Vaud, HES-SO, University of Applied Sciences and Arts Western Switzerland)
Mosqueron, Romuald (School of Engineering and Management Vaud, HES-SO, University of Applied Sciences and Arts Western Switzerland)
Pandey, Om Jee (IIT BHU Varanasi, Varanasi, India)
Reddy Cenkeramaddi, Linga (School of Engineering and Management Vaud, HES-SO, University of Applied Sciences and Arts Western Switzerland)

Date 2022-06

Publié dans IEEE Sensors Journal

Volume 2022, vol. 22, no. 14, pp. 14610-14619

Pagination 10 p.

DOI https://doi.org/10.1109/JSEN.2022.3181518

ISSN 1530-437X

Mots-clés (libres) gesture recognition ; cameras ; convolutional neural networks ; feature extraction ; streaming media ; hidden Markov models ; video sequences ; hand-gestures ; human-computer interaction ; video hand-gestures ; hand-gestures recognition ; RGB-D camera ; light weight CNN

Type d'article scientifique

Domaine Ingénierie et Architecture

Ecole HEIG-VD

Institut ReDS - Reconfigurable & embedded Digital Systems

Le document apparaît dans Articles scientifiques
Global

Résumé

Détails

Actions