Aller au contenu principal
Accès ouvert déclaré 2026 article

Speech Emotion Recognition and Algorithmic Bias

0Citations signalées, ce qui n’est pas une note de qualité
1Institutions déclarées
1Pays d’affiliation déclarés

Rattachement africain : lb. Niveau de preuve : code pays fourni par la source.

Le résumé fourni par la source

Although speech emotion recognition (SER) systems have been integrated into many sectors, including human-computer interaction, health, and educational technologies, concerns about their robustness and fairness persist, particularly across datasets, features, model architecture, and training strategies. This paper explores the impact of different SER design choices, particularly biases, across datasets, features, model architecture, and training strategies. The study uses IEMOCAP and RAVDESS as benchmark corpora. It compares a support vector machine (SVM) with handcrafted acoustic features to a convolutional and recurrent neural network (CRNN) model that operates on a spectrogram, as well as a fine-tuned self-supervised speech encoder. To the best of our knowledge, this is the first study to jointly examine cross-corpus generalization, fairness-aware training, and subgroup disparities by gender and corpus within a unified SER framework. The work explores the effects of fairness-centric training, data augmentation, and class imbalance on performance and subgroup disparity (gender and corpus). The results indicate that,, for deep models, large cross-corpus performance drops (10–30 percentage points) and emotion-specific confusion spersisth is an improvement over the SVM baseline (~65% macro F1) in spein A class-balanced training mechanism and data augmentation are effective in enhancing recognition of minority emotions, raising F1 scores for emotions such as fear and disgust from below 0.40 to 0.52–0.58, while fairness-centric loss mechanisms are effective at reducing the performance gap (from 4–8% to 2–3%) at the cost of global accuracy (1–3% reduction).

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
Speech Emotion Recognition and Algorithmic Bias
Date Crossref
17/09/2026
Éditeur
Bilad Alrafidain University College
Type
journal-article

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude, et il ne compte pas comme une seconde source scientifique indépendante.

Les institutions déclarées

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Les sujets associés

Emotion and Mood RecognitionSentiment Analysis and Opinion MiningSpeech Recognition and Synthesis

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.