Skip to content
Open access

Model Development and Validation for Repetition Severity Assessment in Stuttered Speech Using Clinical Speech Datasets

Aug 2026 · International journal of computer information systems and industrial management applications · 0 citations

TL;DR

A computational model that grades repetition severity from clinical speech recordings, evaluated on 480 audio samples from 60 adult speakers with persistent developmental stuttering, supports its use as a clinical decision-support tool.

Abstract

Stuttering disrupts the forward flow of speech through involuntary repetitions, prolongations, and blocks, with repetitions being the most common and clinically telling of the three. Quantifying how severe those repetitions are matters for diagnosis, therapy planning, and tracking whether treatment is actually working. The catch is that severity has long been judged by ear, by trained speech-language pathologists counting disfluent events and assigning ratings on standardised scales. That process is slow, variable between clinicians, and constrained by who is available. This paper builds and tests a computational model that grades repetition severity from clinical speech recordings. The model draws on a multimodal feature set combining acoustic, prosodic, temporal, and spectral descriptors, evaluated on 480 audio samples from 60 adult speakers with persistent developmental stuttering. Two certified clinicians labelled each sample as mild, moderate, or severe, with strong inter-rater agreement. Recursive feature elimination trimmed the feature set to a compact, discriminative subset, and five classifiers were trained under stratified ten-fold cross-validation repeated five times. A Gradient Boosted Trees model reached 91.46 percent mean accuracy, a macro F1-score of 0.90, and a Cohen kappa of 0.86, ahead of Random Forest, a Support Vector Machine, a Multilayer Perceptron, and Logistic Regression. Per-class scores were strong for mild and severe cases and weaker for moderate, where acoustic characteristics overlap with both neighbouring classes. Friedman and Nemenyi tests confirmed the top model's lead was significant at the 0.05 level. The pipeline is reproducible and the results support its use as a clinical decision-support tool.

Read PDF

Similar papers

Case report Open access Aug 2026

Rapid syllable transition treatment (ReST) in children with speech motor delay: case studies

It is concluded that ReST is effective for children with SMD, including when applied through telepractice, representing a promising evidence-based therapeutic alternative grounded in motor learning principles and supported by linguistically controlled stimuli selection.

Ana Carolina Bucci, Camila Goulart, Aline Mara de Oliveira · 0 citations
#natural language process... Preprint Sep 2026

Reasoning Beyond Transcription: Audio Language Models on Child Stuttering Speech

Results show that while ALMs extract high-level meaning from stuttered speech, reasoning degrades significantly with increased usage, and instruction-guided models are instruction-guided to focus on the child, preserve clinically relevant disfluencies, and avoid adult-speech leakage.

C. Okocha, Christan Grant, Zoey Liu · 1 citation
Conference Aug 2026

AI Based Automatic Speech Assessment and Personalized Therapy for Aphasia Patients

Aphasia is a neurological condition that affects an individual's ability to speak and understand language. Accurate analysis and evaluation of patient speech are essential for effective rehabilitation. In this study, an artificial intelligencebased speech therapy system is proposed to support aphasia rehabilitation. Sp...

S. Devasurithi, M. Madhumitha, G. Nandhini et al. · 0 citations
Aug 2026

Automated language impairment screening in acute stroke using connected speech

This work analyzed brief story retellings from 86 patients with left-hemisphere stroke and derived discrete linguistic features and embeddings with Large Language Models, providing proof of concept for a fast, largely automated discourse screener of acute LI.

L. Pugalenthi, T. Schnur · 0 citations
Aug 2026

ACES: ascertaining diagnosis classification with elicited speech in individuals with heterogeneous, comorbid psychiatric disorders.

It is suggested that a brief, explainable speech-based assessment may be able to identify individuals who need further evaluation for psychiatric disorders and external validation, bias auditing, and deployment studies are warranted to assess clinical impact.

Sunny X. Tang, Jiefei Li, K. Brosch et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.