Recognizing Emotion from Speech using Machine learning and Deep learning

dc.contributor.authorRoy, Tonny
dc.date.accessioned2026-06-13T03:46:15Z
dc.date.available2026-06-13T03:46:15Z
dc.date.issued2025-01-13
dc.descriptionProject report
dc.description.abstractIn the analysis of psychological disorders, behavioral decision making, human machine interaction application speech recognition is plays a essential role. Speech emotion recognition is a system that detects emotions from live audio. people from all over the world utilize words to express their emotions, regardless of their origin. In this project, we focus on using machine learning (ML), which employs a dataset and algorithms to predict or detect any future possibilities. The data sets of audio files in wave format with 8 emotional states: anger, disgust, fear, happiness, pleasant, surprise, sadness, and neutral. Using the librosa library, features were extracted from the audio files in the datasets. The features were applied to multiple machine learning models and results were compared. Speech Emotion Recognition is a popular study topic with numerous applications. It has also became a challenge in the field of speech recognition processing too. Overall, a CNN model would be a good method to human speech emotion recognition with the accuracy rate 85%, because of its capacity to extract complicated patterns and characteristics from input data. The other two models accuracy rates are, SVM 82% and MLP 83%. However, the model's success would be determined by the quality of the preprocessed data, the model architecture used, and the efficacy of the data augmentation strategies employed
dc.identifier.otherhttp://dspace.daffodilvarsity.edu.bd:8080/handle/123456789/17296
dc.identifier.urihttp://dspace.daffodilvarsity.edu.bd:8080/handle/123456789/17296
dc.language.isoen_US
dc.publisherDaffodil International University
dc.sourceDIU Institutional Repository
dc.subjectSpeech Emotion Recognition
dc.subjectMachine Learning
dc.subjectCNN
dc.subjectAudio Feature Extraction
dc.subjectWave Audio Dataset
dc.subjectHuman–Machine Interaction
dc.titleRecognizing Emotion from Speech using Machine learning and Deep learning
dc.typeOther

Files

Original bundle

Now showing 1 - 1 of 1
No Thumbnail Available
Name:
191-15-12650.pdf.txt
Size:
54.98 KB
Format:
Adobe Portable Document Format

Collections