Frag' FlorenceEvidenz. Klar. Anwendbar.
Uhr 7/8Sources Journal Tree
Easy Demo

Crossref · journal-article

Model Development and Validation for Repetition Severity Assessment in Stuttered Speech Using Clinical Speech Datasets

Pooja J N, H Y Vani, Rakshith P, M A Anusuya, Swetha P M

International Journal of Computer Information Systems and Industrial Management Applications · 2026 · Band 18 · Ausgabe 16s · S. 299-313

Vollständiger Abstract

Worum geht es in dieser Arbeit?

Stuttering disrupts the forward flow of speech through involuntary repetitions, prolongations, and blocks, with repetitions being the most common and clinically telling of the three. Quantifying how severe those repetitions are matters for diagnosis, therapy planning, and tracking whether treatment is actually working. The catch is that severity has long been judged by ear, by trained speech-language pathologists counting disfluent events and assigning ratings on standardised scales. That process is slow, variable between clinicians, and constrained by who is available. This paper builds and tests a computational model that grades repetition severity from clinical speech recordings. The model draws on a multimodal feature set combining acoustic, prosodic, temporal, and spectral descriptors, evaluated on 480 audio samples from 60 adult speakers with persistent developmental stuttering. Two certified clinicians labelled each sample as mild, moderate, or severe, with strong inter-rater agreement. Recursive feature elimination trimmed the feature set to a compact, discriminative subset, and five classifiers were trained under stratified ten-fold cross-validation repeated five times. A Gradient Boosted Trees model reached 91.46 percent mean accuracy, a macro F1-score of 0.90, and a Cohen kappa of 0.86, ahead of Random Forest, a Support Vector Machine, a Multilayer Perceptron, and Logistic Regression. Per-class scores were strong for mild and severe cases and weaker for moderate, where acoustic characteristics overlap with both neighbouring classes. Friedman and Nemenyi tests confirmed the top model's lead was significant at the 0.05 level. The pipeline is reproducible and the results support its use as a clinical decision-support tool.

Bibliografischer Nachweis

Publikationsdaten

Autor:innen
Pooja J N, H Y Vani, Rakshith P, M A Anusuya, Swetha P M
Quelle
International Journal of Computer Information Systems and Industrial Management Applications
Publikation
2026-08-12
Band / Ausgabe
18 / 16s
Seiten
299-313
ISSN / ISBN
2150-7988, 2150-7988
Zitationen
0 laut Crossref
Referenzen
0 hinterlegt

Zitieren

Zitierfähiger Nachweis

Pooja J N, H Y Vani, Rakshith P, M A Anusuya, Swetha P M (2026). Model Development and Validation for Repetition Severity Assessment in Stuttered Speech Using Clinical Speech Datasets. International Journal of Computer Information Systems and Industrial Management Applications, 18 (16s), 299-313. https://doi.org/10.70917/ijcisim-2026-4576
RIS BibTeX CSL-JSON