Ensemble Learning for Speech Emotion Recognition using Graph-Based Signal Dynamics

سال انتشار: 1405
نوع سند: مقاله ژورنالی
زبان: انگلیسی
مشاهده: 107

فایل این مقاله در 12 صفحه با فرمت PDF قابل دریافت می باشد

استخراج به نرم افزارهای پژوهشی:

لینک ثابت به این مقاله:

شناسه ملی سند علمی:

JR_JADM-14-2_006

تاریخ نمایه سازی: 26 فروردین 1405

چکیده مقاله:

Nowadays, the recognition of emotions using speech signals has gained popularity because of its vast number of applications in different fields such as medicine, online marketing, online search engines, education systems, criminal investigations, traffic collisions, and more. Many researchers have adopted different methodologies to improve emotion classification accuracy using speech signals. This study presents a novel time-series-to-graph transformation framework for speech emotion recognition. Speech signals were segmented into overlapping windows, each converted into graphs, from which ۱۶ structural features were extracted. Significant features were then selected via Minimum Redundancy Maximum Relevance (mRMR) and used to train four classifiers: random forest (RF), linear discriminant analysis (LDA), support vector machine (SVM), and k-nearest neighbors (KNN). Finally, a soft-voting ensemble strategy was employed to integrate their predictions, yielding improved classification performance. The proposed method achieved the highest sensitivity, specificity, and accuracy for the SAVEE database: ۸۳.۵۷%, ۹۸.۹۳%, and ۹۸.۱۶%, respectively. Similarly, for the EmoDB database, the highest values were ۹۴.۴۷%, ۹۹.۰۹%, and ۹۸.۴۰%, respectively. We also compared our results with other methods and found that our method outperformed state-of-the-art techniques in emotion classification.

نویسندگان

Zeynab Mohammadpoory

Faculty of Electrical Engineering, Shahrood University of Technology, Shahrood, Iran.

Mahda Nasrollahzadeh

Department of Electrical Engineering, University of Torbat Heydarieh, Torbat Heydarieh, Iran.

Sakineh Asadi

Department of Computer Engineering, University of Mazandaran, Babolsar, Iran.

مراجع و منابع این مقاله:

لیست زیر مراجع و منابع استفاده شده در این مقاله را نمایش می دهد. این مراجع به صورت کاملا ماشینی و بر اساس هوش مصنوعی استخراج شده اند و لذا ممکن است دارای اشکالاتی باشند که به مرور زمان دقت استخراج این محتوا افزایش می یابد. مراجعی که مقالات مربوط به آنها در سیویلیکا نمایه شده و پیدا شده اند، به خود مقاله لینک شده اند :
  • M. Wang, H. Ma, Y. Wang, and X. Sun, "Design ...
  • S. P. Mishra, P. Warule, and S. Deb, "Speech emotion ...
  • H. Wang, Y. Liu, X. Zhen, and X. Tu, "Depression ...
  • M. Bojanić, V. Delić, and A. Karpov, "Call redistribution for ...
  • T. Deschamps-Berger, L. Lamel, and L. Devillers, "End-to-end speech emotion ...
  • X. Cai, D. Dai, Z. Wu, X. Li, J. Li, ...
  • M. El Ayadi, M. S. Kamel, and F. Karray, "Survey ...
  • J. Kacur, B. Puterka, J. Pavlovicova, and M. Oravec, "On ...
  • Z. Mohammadpoory, M. Nasrolahzadeh, S. A. Amiri, and J. Haddadnia, ...
  • T. M. Wani, T. S. Gunawan, S. A. A. Qadri, ...
  • M. Nasrolahzadeh, Z. Mohammadpoory, and J. Haddadnia, "Weighted Visibility Graph-based ...
  • F. Mohammady, S. Asadi Amiri, and Z. Mohammadpoory, "Leveraging segmentation ...
  • M. Nasrolahzadeh, Z. Mohammadpoory, and J. Haddadnia, "Indices from visibility ...
  • L. Lacasa, B. Luque, J. Luque, and J. C. Nuno, ...
  • Z. Mohammadpoory, M. Nasrolahzadeh, S. A. Amiri, "Classification of healthy ...
  • M. Nasrolahzadeh, Z. Mohammadpoory, and J. Haddadnia, "The visibility graph ...
  • Y. You, C. Cai, and Y. Wu, "۳D visibility graph ...
  • T. Varoudis and S. Psarra, "Beyond two dimensions: architecture through ...
  • Z. Mohammadpoory, M. Nasrolahzadeh, S. A. Amiri, "Patient-independent epileptic seizure ...
  • M. Nasrolahzadeh, Z. Mohammadpoory, and J. Haddadnia, "A novel method ...
  • E. Lieskovská, M. Jakubec, R. Jarina, and M. Chmulík, "A ...
  • B. Schuller and A. Batliner, Computational Paralinguistics: Emotion, Affect and ...
  • M. Papakostas, G. Siantikos, T. Giannakopoulos, E. Spyrou, and D. ...
  • T. M. Wani, T. S. Gunawan, S. A. A. Qadri, ...
  • K. Tomba, J. Dumoulin, E. Mugellini, O. Abou Khaled, and ...
  • L. Vignolo, H. Rufiner, and D. Milone, "Multi-objective optimisation of ...
  • K. Aghajani and I. E. Paeen Afrakoti, "Speech emotion recognition ...
  • M. El Ayadi, M. S. Kamel, and F. Karray, "Survey ...
  • S. Taran, "A nonlinear feature extraction approach for speech emotion ...
  • R. K. Srivastava and D. Pandey, "Speech recognition using HMM ...
  • J. M. López-Gil and N. Garay-Vitoria, "Assessing the effectiveness of ...
  • Y. Pan, P. Shen, and L. Shen, "Speech emotion recognition ...
  • P. Shegokar and P. Sircar, "Continuous wavelet transform based speech ...
  • S. S. Poorna, V. Menon, and S. Gopalan, "Hybrid CNN-BiLSTM ...
  • M. Nasrolahzadeh, Z. Mohammadpoory, "A novel method for distinction heart ...
  • M. Ahmadlou, H. Adeli, and A. Adeli, "New diagnostic EEG ...
  • Z. Mohammadpoory, M. Nasrolahzadeh, and J. Haddadnia, "Epileptic seizure detection ...
  • Z. Mohammadpoory, M. Nasrolahzadeh, N. Mahmoodian, M. Sayyah, and J. ...
  • M. Nasrolahzadeh, Z. Mohammadpoory, and J. Haddadnia, "Analysis of heart ...
  • Z. Mohammadpoory, M. Nasrolahzadeh, N. Mahmoodian, and J. Haddadnia, "Automatic ...
  • S. Haq and P. J. B. Jackson, "Speaker-dependent audio-visual emotion ...
  • F. Burkhardt, A. Paeschke, M. Rolfes, W. F. Sendlmeier, and ...
  • S. Asadi Amiri, M. Nasrolahzadeh, Z. Mohammadpoory, A. Movahedinia, and ...
  • R. A. Fisher, "The use of multiple measurements in taxonomic ...
  • L. Breiman, "Bagging predictors," Mach. Learn., vol. ۲۴, no. ۲, ...
  • N. V. Chawla, K. W. Bowyer, L. O. Hall, and ...
  • S. Moghani, H. Marvi, and Z. Mohammadpoory, “Valvular Heart Disease ...
  • S. P. Mishra, P. Warule, and S. Deb, "Chirplet transform ...
  • J. Xie, M. Zhu, and K. Hu, "Fusion-based speech emotion ...
  • S. P. Mishra, P. Warule, and S. Deb, "Speech emotion ...
  • نمایش کامل مراجع