Aller au contenu principal
Accès ouvert déclaré 2026 preprint

Automatic Speech Recognition and Phonetics-Informed Sentence Design for Spastic Dysarthria Detection and Corticobulbar Lesion Localization

0Citations signalées, ce qui n’est pas une note de qualité
3Institutions déclarées
1Pays d’affiliation déclarés

Rattachement africain : th. Niveau de preuve : code pays fourni par la source.

Le résumé fourni par la source

Abstract Spastic dysarthria diagnosis through subjective neurologist auditory-perceptual assessment remains standard practice despite known inaccuracy. To address this gap, we developed an objective framework grounded in phonetic evidence that spastic dysarthria preferentially impairs initial consonant articulation, using automatic speech recognition (ASR) to quantify dysarthria and localize corticobulbar lesions. We created four reading sentences targeting groups of initial consonants: labial (facial), lingual-alveolar (tongue), and velopharyngeal (pharyngeal/soft-palate) sentence, along with a mixed-consonant sentence for comparative evaluation. Thirty-seven patients with neuroimaging-confirmed corticobulbar lesions and 37 controls read each sentence. ASR transcribed dysarthric speech into text, and we computed a ‘syllable-error score’ by counting incorrectly transcribed syllables. This yields a clinically meaningful feature that makes syllable-level phonetic errors explicit. Logistic regression models were trained for each sentence, and performance was summarized by the area under the receiver operating characteristic curve (AUC) across 10,000 resampled train-test splits. Consonant-specific sentences significantly outperformed the mixed sentence: the lingual-alveolar sentence performed best with (median AUC 0.88), followed by the labial (0.80), then the velopharyngeal sentence (0.72), while the mixed-consonant sentence was lowest (0.67). These results suggest that the interpretable ASR-derived syllable error feature, combined with a relevant machine learning classifier could inform clinical insight into consonant-specific vulnerability in spastic dysarthria, with lingual-alveolar consonants appearing particularly informative. Overall, this novel ASR-based framework, together with phonetics-informed feature design provides objective, accurate, and clinically meaningful digital quantification for spastic dysarthria detection and corticobulbar lesion localization. Author summary At the bedside, neurologists often detect dysarthria by listening to a patient’s speech, a practical but subjective approach that may delay diagnosis when dysarthria is mild. We developed a simple digital assessment approach that combines clinical phonetic knowledge with an accessible artificial intelligence tool, automatic speech recognition (ASR), to make the bedside judgement more objective while remaining clinically meaningful. Rather than relying on pre-existing open speech datasets, we newly collected speech recordings from patients with neuroimaging-confirmed corticobulbar lesions and matched healthy participants. Most patients had mild dysarthria, representing a setting where perceptual diagnosis is uncertain. We designed sentences that target different speech muscle groups and used ASR transcription errors as measurable features. A simple machine learning classifier was then trained on these features to evaluate diagnostic performance. The best sentence designs distinguished patients from controls with good performance and produced understandable results in relation to speech physiology. Our study illustrates a broader principle for digital neurology: artificial intelligence may be most useful when it is guided by clinical knowledge rather than replacing it and when it is available at the bedside. This approach could shift neurological assessment toward objective yet explainable way and could be extended beyond diagnosis to repeated monitoring during speech rehabilitation.

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
Automatic Speech Recognition and Phonetics-Informed Sentence Design for Spastic Dysarthria Detection and Corticobulbar Lesion Localization
Date Crossref
03/06/2026
Éditeur
openRxiv
Type
posted-content

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude, et il ne compte pas comme une seconde source scientifique indépendante.

Les institutions déclarées

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Les sujets associés

Voice and Speech DisordersDysphagia Assessment and ManagementStuttering Research and Treatment

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.