Speech Emotion Recognition Using Deep Learning
摘要
One of the challenges of developing a machine learning-based system for recognising emotions in speech signals is the variability of human emotion expression. Depending on the situation and the individual, the same emotion can be conveyed in a variety of ways by various people. As a result, it is crucial to compile a broad selection of annotated speech recordings that showcase a variety of expressions and emotional states. The system should also be built to deal with clamorous settings and various speaking styles, including accents and speech difficulties. We can harness the potential of emotion recognition in speech signals to enhance human–computer interaction and help with mental health diagnosis and treatment by solving these issues and creating a strong system.