Skip to main navigation Skip to search Skip to main content

A hierarchical model for learning to understand head gesture videos

  • Shandong University
  • Key Lab of the Ministry of Education for Process Control and Efficiency Egineering
  • University of South Carolina

Research output: Contribution to journalArticlepeer-review

5 Scopus citations

Abstract

Head gesture videos recorded of a person bear rich information about the individual. Automatically understanding these videos can empower many useful human-centered applications in areas such as smart health, education, work safety and security. To understand a video's content, low-level head gesture signals carried in the video that capture characteristics of both human postures and motions need to be translated into high-level semantic labels. To meet this aim, we propose a hierarchical model for learning to understand head gesture videos. Given a head gesture video of an arbitrary length, the model first segments the full-length video into multiple short clips for clip-based feature extraction. Multiple base feature extraction procedures are then independently tuned via a set of peripheral learning tasks without consuming any labels of the goal task. These independently derived base features are subsequently aggregated through a multi-task learning framework, coupled with a feature dimensionality reduction module, to optimally learn to accomplish the end video understanding task in an weakly supervised manner, utilizing the limited amount of video labels available of the goal task. Experimental results show that the hierarchical model is superior to multiple state-of-the-art peer methods in tackling versatile video understanding tasks.

Original languageEnglish
Article number108256
JournalPattern Recognition
Volume121
DOIs
StatePublished - Jan 2022
Externally publishedYes

UN SDGs

This output contributes to the following UN Sustainable Development Goals (SDGs)

  1. SDG 3 - Good Health and Well-being
    SDG 3 Good Health and Well-being

Keywords

  • Head gesture videos
  • Multi-task learning
  • Stacked BLSTM
  • Transfer learning
  • Video understanding

Fingerprint

Dive into the research topics of 'A hierarchical model for learning to understand head gesture videos'. Together they form a unique fingerprint.

Cite this