Privacy preserving federated learning approach for speech emotion recognition
Loading...
Date
Publisher
Institute of Electrical and Electronics Engineers Inc.
Citation
M. R. Z. Chowdhury et al., "Privacy Preserving Federated Learning Approach for Speech Emotion Recognition," 2023 26th International Conference on Computer and Information Technology (ICCIT), Cox's Bazar, Bangladesh, 2023, pp. 1-6, doi: 10.1109/ICCIT60459.2023.10441577.
Abstract
Emotions are a critical factor in intrapersonal communication and significantly influence how we convey our intentions and feelings. Recognizing emotion through speech not only enhances our understanding of interpersonal dynamics but also holds immense potential across diverse sectors such as healthcare, human-machine interaction, automated customer service and more. However, the majority of the existing speech recognition systems are highly centralized, raising concerns over potential data leakage. To address this issue, we have introduced a privacy preserving system to recognize emotion from audio data using federated learning. Our approach leveraged the distributed model training and aggregation strategy, ensuring data privacy while eliminating the need for data sharing to a centralized system. Moreover, we explored the potential of CNN and LSTM models, both as a distributed and centralized formats, with MFCC as features in the federated learning setting. Experimental evaluation of our approach on the IEMOCAP and CREMA-D datasets achieved a maximum accuracy of 68.65% and 68.82%, surpassing the existing federated learning techniques and rivaling centralized benchmarks.
LC Subject Headings
Description
Publisher Link
Type
Conference Proceeding