Frag' FlorenceEvidenz. Klar. Anwendbar.
Uhr 7/8Sources Journal Tree
Easy Demo

Lokaler Crossref-Datenbestand · journal-article

Artificial intelligence responses to toilet training questions: A reliability and readability analysis of ChatGPT and gemini applications

Ayşegül Sarıkaya, Ayşe Alptekin

WORK: A Journal of Prevention, Assessment & Rehabilitation · 2026

Vollständiger Abstract

Worum geht es in dieser Arbeit?

Background Toilet training is a critical developmental milestone that may have long-term implications for children's psychosocial development if improperly guided. With the increasing use of artificial intelligence (AI) chatbot applications for health-related information, evaluating the reliability and readability of AI-generated guidance on sensitive developmental topics has become essential. Objective This study aimed to evaluate the reliability and readability of responses provided by AI chatbot applications regarding commonly asked questions about toilet training. Methods Two widely used AI platforms—ChatGPT-4 Turbo (OpenAI) and Gemini 2.0 Flash (Google)—were included in the study. The study was initiated on April 29, 2025, by submitting a standardized prompt to the chatbots. Ten frequently asked questions about toilet training were selected based on AI-generated query lists and Google Trends data. Responses were obtained in independent sessions and evaluated by a panel of five child development experts using a four-point Likert-type scale developed by Mika et al. Readability levels were assessed using the Flesch-Kincaid Grade Level through WordCalc software. Statistical analyses were conducted to compare quality and readability across platforms. Results Statistically significant differences in response quality were identified for Questions 2, 3, and 7 (p < 0.05). In the quality rating system, lower scores indicate higher response quality. Gemini demonstrated lower (better) median quality scores for these items. Regarding readability, Gemini produced responses with a higher Flesch–Kincaid Grade Level (i.e., more complex reading level), particularly for Questions 3 and 7. No statistically significant differences in response quality were found for the remaining seven questions. Conclusions Both AI applications provided generally acceptable expert-rated responses to common toilet-training questions; however, differences in response quality and readability were observed across specific items. These findings suggest that AI tools may serve as accessible supplementary informational resources for families, but they should not be interpreted as substitutes for professional guidance or as evidence of clinical validity.

Bibliografischer Nachweis

Publikationsdaten

Autor:innen
Ayşegül Sarıkaya, Ayşe Alptekin
Quelle
WORK: A Journal of Prevention, Assessment & Rehabilitation
Publikation
2026-01-01
Band / Ausgabe
Nicht angegeben
Seiten
Nicht angegeben
ISSN / ISBN
1051-9815, 1875-9270
Zitationen
0 laut Crossref
Referenzen
0 hinterlegt

Zitieren

Zitierfähiger Nachweis

Ayşegül Sarıkaya, Ayşe Alptekin (2026). Artificial intelligence responses to toilet training questions: A reliability and readability analysis of ChatGPT and gemini applications. WORK: A Journal of Prevention, Assessment & Rehabilitation. https://doi.org/10.1177/10519815261477742
RIS BibTeX CSL-JSON

Kontext

Themen, Förderung und Nutzung

Lizenzhinweise: Lizenz 1