Research on facial expression recognition algorithm based on improved MobileNetV3

Bin Jiang; Nanxing Li; Xiaomei Cui; Qiuwen Zhang; Huanlong Zhang; Zuhe Li; Weihua Liu

doi:10.1186/s13640-024-00638-z

EURASIP Journal on Image and Video Processing (Aug 2024)

Research on facial expression recognition algorithm based on improved MobileNetV3

Bin Jiang,
Nanxing Li,
Xiaomei Cui,
Qiuwen Zhang,
Huanlong Zhang,
Zuhe Li,
Weihua Liu

Affiliations

Bin Jiang: College of Computer and Communication Engineering, Zhengzhou University of Light Industry
Nanxing Li: College of Computer and Communication Engineering, Zhengzhou University of Light Industry
Xiaomei Cui: College of Computer and Communication Engineering, Zhengzhou University of Light Industry
Qiuwen Zhang: College of Computer and Communication Engineering, Zhengzhou University of Light Industry
Huanlong Zhang: College of Electric and Information Engineering, Zhengzhou University of Light Industry
Zuhe Li: College of Computer and Communication Engineering, Zhengzhou University of Light Industry
Weihua Liu: College of Computer and Communication Engineering, Zhengzhou University of Light Industry

DOI: https://doi.org/10.1186/s13640-024-00638-z
Journal volume & issue: Vol. 2024, no. 1
pp. 1 – 16

Abstract

Read online

Abstract Aiming at the problem that face images are easily interfered by occlusion factors in uncontrollable environments, and the complex structure of traditional convolutional neural networks leads to low expression recognition rates, slow network convergence speed, and long network training time, an improved lightweight convolutional neural network is proposed for facial expression recognition algorithm. First, the dilation convolution is introduced into the shortcut connection of the inverted residual structure in the MobileNetV3 network to expand the receptive field of the convolution kernel and reduce the loss of expression features. Then, the channel attention mechanism SENet in the network is replaced by the two-dimensional (channel and spatial) attention mechanism SimAM introduced without parameters to reduce the network parameters. Finally, in the normalization operation, the Batch Normalization of the backbone network is replaced with Group Normalization, which is stable at various batch sizes, to reduce errors caused by processing small batches of data. Experimental results on RaFD, FER2013, and FER2013Plus face expression data sets show that the network reduces the training times while maintaining network accuracy, improves network convergence speed, and has good convergence effects.

Published in EURASIP Journal on Image and Video Processing

ISSN: 1687-5176 (Print); 1687-5281 (Online)
Publisher: SpringerOpen
Country of publisher: United Kingdom
LCC subjects: Technology: Electrical engineering. Electronics. Nuclear engineering: Electronics
Website: https://jivp-eurasipjournals.springeropen.com

About the journal

Abstract

Keywords