Clinical Research Directory
Browse clinical research sites, groups, and studies.
Large Language Models for Stuttering Assessment and Therapy
Sponsor: Istanbul Gelisim University
Summary
This observational study aims to compare the quality of responses generated by four large language models (ChatGPT, Claude, Google Gemini, and Grok-4.20) to questions about stuttering. The study focuses on three main areas: general information about stuttering, clinical assessment, and therapy. A total of nine questions were developed based on evidence-based clinical practice guidance, including the American Speech-Language-Hearing Association (ASHA) Practice Portal. Each question is presented to each language model in separate sessions, and the generated responses are recorded for evaluation. Five speech-language therapists with clinical experience in stuttering independently evaluate the model-generated responses. Each response is rated for relevance, accuracy, clarity, completeness, and consistency using a 5-point Likert scale. The responses are also compared with guideline-based reference information. The study does not involve any clinical intervention or treatment of patients. The aim is to determine how closely large language model responses align with current clinical practice guidance and to identify their potential strengths and limitations when used as informational or clinical support tools in the field of stuttering.
Official title: Alignment of Large Language Model Responses on Stuttering Definition, Assessment, and Therapy With Clinical Practice Guidelines: A Multi-Model Comparison
Key Details
Gender
All
Age Range
18 Years - Any
Study Type
OBSERVATIONAL
Enrollment
5
Start Date
2026-08-07
Completion Date
2026-09-06
Last Updated
2026-09-22
Healthy Volunteers
Yes
Conditions
Interventions
ChatGPT-Generated Responses
ChatGPT responses to nine standardized questions on stuttering information, assessment, and therapy were generated in separate sessions and recorded for independent expert evaluation. Questions were repeated under comparable conditions to allow assessment of response consistency.
Claude-Generated Responses
Claude responses to nine standardized questions on stuttering information, assessment, and therapy were generated in separate sessions and recorded for independent expert evaluation. Questions were repeated under comparable conditions to allow assessment of response consistency.
Google Gemini-Generated Responses
Google Gemini responses to nine standardized questions on stuttering information, assessment, and therapy were generated in separate sessions and recorded for independent expert evaluation. Questions were repeated under comparable conditions to allow assessment of response consistency.
Grok-4.20-Generated Responses
Grok-4.20 responses to nine standardized questions on stuttering information, assessment, and therapy were generated in separate sessions and recorded for independent expert evaluation. Questions were repeated under comparable conditions to allow assessment of response consistency.
Locations (1)
Istanbul Gelisim University
Istanbul, Istanbul, Turkey (Türkiye)