Vollständiger Abstract
Worum geht es in dieser Arbeit?
Background: Psychological assessment is essential for bariatric surgery candidacy but remains inconsistent and time-consuming. This study compared three large language models (LLMs) in generating structured psychological screening checklists to support standardized preoperative evaluation. Methods: Models A, B, and C generated checklists for five standardized bariatric vignettes. Three blinded psychologists independently rated the outputs on 5-point Likert scales for clinical relevance, completeness, specificity, and usability. Statistical analysis included Jaccard similarity to quantify content overlap and analysis of variance with Tukey post hoc tests to compare expert ratings. Results: Model A produced the most extensive checklists (mean = 35.6 items, standard deviation = 4.2), while Models B and C were more concise (28.4 and 26.7 items, respectively). Model A achieved the highest completeness scores, whereas Model B was rated highest for specificity and usability. Interrater agreement was good to excellent (intraclass correlation coefficient = 0.79–0.87). Moderate content overlap (Jaccard similarity = 0.45–0.63) suggested complementary model strengths. Between-model differences were significant for completeness, F (2, 6) = 11.5, p = 0.008, and specificity, F (2, 6) = 9.8, p = 0.013. Conclusions: LLMs differ significantly in checklist breadth and focus. Strategic model selection can enhance preoperative workflows by balancing comprehensiveness with targeted, actionable assessment. However, deployment should remain clinician-supervised because outputs may vary with model updates, prompt phrasing, and local workflow constraints.
Bibliografischer Nachweis
Publikationsdaten
- Autor:innen
- Yahya Kemal Çalışkan, Fatih Başak
- Quelle
- Bariatric Surgical Practice and Patient Care
- Publikation
- 2026-01-01
- Band / Ausgabe
- Nicht angegeben
- Seiten
- Nicht angegeben
- ISSN / ISBN
- 2168-023X, 2168-0248
- Zitationen
- 0 laut Crossref
- Referenzen
- 0 hinterlegt
Zitieren
Zitierfähiger Nachweis
Yahya Kemal Çalışkan, Fatih Başak (2026). Comparative Performance and Utility of Large Language Models in Generating Psychological Screening Checklists for Bariatric Surgery. Bariatric Surgical Practice and Patient Care. https://doi.org/10.1177/2168023x261481848
Kontext
Themen, Förderung und Nutzung
Lizenzhinweise: Lizenz 1