Improving real-time hand gesture recognition with semantic segmentation

Gibran Benitez-Garcia, Lidia Prudente-Tixteco, Luis Carlos Castro-Madrid, Rocio Toscano-Medina, Jesus Olivares-Mercado, Gabriel Sanchez-Perez, Luis Javier Garcia Villalba

Research output: Contribution to journalArticlepeer-review

37 Scopus citations

Abstract

Hand gesture recognition (HGR) takes a central role in human–computer interaction, cov-ering a wide range of applications in the automotive sector, consumer electronics, home automation, and others. In recent years, accurate and efficient deep learning models have been proposed for real-time applications. However, the most accurate approaches tend to employ multiple modalities derived from RGB input frames, such as optical flow. This practice limits real-time performance due to intense extra computational cost. In this paper, we avoid the optical flow computation by proposing a real-time hand gesture recognition method based on RGB frames combined with hand segmentation masks. We employ a light-weight semantic segmentation method (FASSD-Net) to boost the accuracy of two efficient HGR methods: Temporal Segment Networks (TSN) and Temporal Shift Modules (TSM). We demonstrate the efficiency of the proposal on our IPN Hand dataset, which includes thirteen different gestures focused on interaction with touchless screens. The experimental results show that our approach significantly overcomes the accuracy of the original TSN and TSM algorithms by keeping real-time performance.

Original languageEnglish
Article number356
Pages (from-to)1-16
Number of pages16
JournalSensors (Switzerland)
Volume21
Issue number2
DOIs
StatePublished - 2 Jan 2021

Keywords

  • FASSD-Net
  • Hand gesture recognition
  • Hand segmentation
  • TSM
  • TSN

Fingerprint

Dive into the research topics of 'Improving real-time hand gesture recognition with semantic segmentation'. Together they form a unique fingerprint.

Cite this