Articulatory-to-Acoustic Conversion of Mandarin Emotional Speech Based on PSO-LSSVM

Guofeng Ren; Jianmei Fu; Guicheng Shao; Yanqin Xun

doi:10.1155/2021/8876005

Complexity (Jan 2021)

Articulatory-to-Acoustic Conversion of Mandarin Emotional Speech Based on PSO-LSSVM

Guofeng Ren,
Jianmei Fu,
Guicheng Shao,
Yanqin Xun

Affiliations

Guofeng Ren: Department of Electronics, Xinzhou Teachers University, Xinzhou 034000, China
Jianmei Fu: Department of Electronics, Xinzhou Teachers University, Xinzhou 034000, China
Guicheng Shao: Department of Electronics, Xinzhou Teachers University, Xinzhou 034000, China
Yanqin Xun: Department of Electronics, Xinzhou Teachers University, Xinzhou 034000, China

DOI: https://doi.org/10.1155/2021/8876005
Journal volume & issue: Vol. 2021

Abstract

Read online

The production of emotional speech is determined by the movement of the speaker’s tongue, lips, and jaw. In order to combine articulatory data and acoustic data of speakers, articulatory-to-acoustic conversion of emotional speech has been studied. In this paper, parameters of LSSVM model have been optimized using the PSO method, and the optimized PSO-LSSVM model was applied to the articulatory-to-acoustic conversion. The root mean square error (RMSE) and mean Mel-cepstral distortion (MMCD) have been used to evaluate the results of conversion; the evaluated result illustrates that MMCD of MFCC is 1.508 dB, and RMSE of the second formant (F2) is 25.10 Hz. The results of this research can be further applied to the feature fusion of emotion speech recognition to improve the accuracy of emotion recognition.

Published in Complexity

ISSN: 1076-2787 (Print); 1099-0526 (Online)
Publisher: Wiley
Country of publisher: United Kingdom
LCC subjects: Science: Mathematics: Instruments and machines: Electronic computers. Computer science
Website: https://onlinelibrary.wiley.com/journal/8503

About the journal