Vollständiger Abstract
Worum geht es in dieser Arbeit?
<h4>Objective</h4>This study investigated physicians' use of large language model (LLM)-generated assessments in psychiatric symptom evaluation and factors influencing their acceptance of artificial intelligence (AI) suggestions.<h4>Methods</h4>Twelve psychiatrists evaluated 50 anonymized psychiatric case vignettes, rating the presence and severity of 73 symptoms using a 4-point Likert scale. After initial assessment, they reviewed GPT-4o's symptom ratings and could revise their evaluations. Accuracy was measured against a gold standard established by expert consensus. We computed initial (I), LLM (L), and revised (R) accuracy scores using percent agreement and weighted kappa. Performance improvement was measured by relative kappa increase. Regression analyses examined the influence of task difficulty and user competence on accuracy improvement.<h4>Results</h4>LLM assessments outperformed initial physician ratings in both agreement (90.4% vs. 82.9%, p<0.001) and kappa (0.713 vs. 0.559, p<0.001). After revision, physician performance improved (κ=0.651) but remained below LLM levels. The switch rate-cases where physicians revised in response to LLM disagreement-was modest (25.4%), indicating partial reliance on AI. Performance gains were positively associated with LLM accuracy and negatively associated with users' own baseline competence, suggesting that less confident users rely more on LLM assistance.<h4>Conclusion</h4>Physicians demonstrated limited but strategic trust in LLM outputs, adjusting their judgments more when the LLM was accurate or when their own competence was lower. Miscalibrated trust-excessive skepticism or overreliance-led to missed gains. Effective human-AI collaboration in psychiatry requires tools and training for accurate self-assessment and AI trust calibration to avoid the pitfalls of algorithm aversion or overreliance.
Abstract: PubMed · Datensatz
Bibliografischer Nachweis
Publikationsdaten
- Autor:innen
- B. Esposito, A. Marini, G. Piano-Mortari, F. Ronga, M. Nigro, L. Pescara, R. Bernabei, S. d'Angelo, P. Monacelli, M. Moricca, L. Paoluzi, R. Santonico, F. Sebastiani
- Quelle
- Lettere Al Nuovo Cimento Series 2
- Publikation
- 1981-01-01
- Band / Ausgabe
- Nicht angegeben
- Seiten
- Nicht angegeben
- ISSN / ISBN
- 1827-613X
- Zitationen
- 9 laut Crossref
- Referenzen
- 0 hinterlegt
Zitieren
Zitierfähiger Nachweis
B. Esposito, A. Marini, G. Piano-Mortari, F. Ronga, M. Nigro, L. Pescara, R. Bernabei, S. d'Angelo, P. Monacelli, M. Moricca, L. Paoluzi, R. Santonico, F. Sebastiani (1981). Measurement on $$\pi ^ + \pi ^ - \pi ^0 \pi ^0 $$ , $$\pi ^ + \pi ^ - \pi ^ + \pi ^ - \pi ^0 $$ , $$\pi ^ + \pi ^ - \pi ^ + \pi ^ - \pi ^0 \pi ^0 $$ , $$\pi ^ + \pi ^ - \pi ^ + \pi ^ - \pi ^ + \pi ^ - $$ production cross-sections in e+e− annihilation at (1.45÷1.80) GeV c.m. Energyproduction cross-sections in e+e− annihilation at (1.45÷1.80) GeV c.m. Energy. Lettere Al Nuovo Cimento Series 2. https://doi.org/10.30773/pi.2025.0330
Kontext
Themen, Förderung und Nutzung
Lizenzhinweise: Lizenz 1