Towards Efficient Multi-Modal Emotion Recognition

Simon Dobrišek; Rok Gajšek; France Mihelič; Nikola Pavešić; Vitomir Štruc

doi:10.5772/54002

International Journal of Advanced Robotic Systems (Jan 2013)

Towards Efficient Multi-Modal Emotion Recognition

Simon Dobrišek,
Rok Gajšek,
France Mihelič,
Nikola Pavešić,
Vitomir Štruc

Affiliations

Simon Dobrišek: Faculty of Electrical Engineering, University of Ljubljana, Ljubljana, Slovenia
Rok Gajšek: Faculty of Electrical Engineering, University of Ljubljana, Ljubljana, Slovenia
France Mihelič: Faculty of Electrical Engineering, University of Ljubljana, Ljubljana, Slovenia
Nikola Pavešić: Faculty of Electrical Engineering, University of Ljubljana, Ljubljana, Slovenia
Vitomir Štruc: Faculty of Electrical Engineering, University of Ljubljana, Ljubljana, Slovenia

DOI: https://doi.org/10.5772/54002
Journal volume & issue: Vol. 10

Abstract

Read online

The paper presents a multi-modal emotion recognition system exploiting audio and video (i.e., facial expression) information. The system first processes both sources of information individually to produce corresponding matching scores and then combines the computed matching scores to obtain a classification decision. For the video part of the system, a novel approach to emotion recognition, relying on image-set matching, is developed. The proposed approach avoids the need for detecting and tracking specific facial landmarks throughout the given video sequence, which represents a common source of error in video-based emotion recognition systems, and, therefore, adds robustness to the video processing chain. The audio part of the system, on the other hand, relies on utterance-specific Gaussian Mixture Models (GMMs) adapted from a Universal Background Model (UBM) via the maximum a posteriori probability (MAP) estimation. It improves upon the standard UBM-MAP procedure by exploiting gender information when building the utterance-specific GMMs, thus ensuring enhanced emotion recognition performance. Both the uni-modal parts as well as the combined system are assessed on the challenging multi-modal eNTERFACE'05 corpus with highly encouraging results. The developed system represents a feasible solution to emotion recognition that can easily be integrated into various systems, such as humanoid robots, smart surveillance systems and alike.

Published in International Journal of Advanced Robotic Systems

ISSN: 1729-8814 (Online)
Publisher: SAGE Publishing
Country of publisher: United Kingdom
LCC subjects: Technology: Electrical engineering. Electronics. Nuclear engineering: Electronics; Science: Mathematics: Instruments and machines: Electronic computers. Computer science
Website: https://journals.sagepub.com/home/arx

About the journal