Frag' FlorenceEvidenz. Klar. Anwendbar.
Uhr 7/8Sources Journal Tree
Easy Demo

Lokaler Crossref-Datenbestand · journal-article

Measurement on $$\pi ^ + \pi ^ - \pi ^0 \pi ^0 $$ , $$\pi ^ + \pi ^ - \pi ^ + \pi ^ - \pi ^0 $$ , $$\pi ^ + \pi ^ - \pi ^ + \pi ^ - \pi ^0 \pi ^0 $$ , $$\pi ^ + \pi ^ - \pi ^ + \pi ^ - \pi ^ + \pi ^ - $$ production cross-sections in e+e− annihilation at (1.45÷1.80) GeV c.m. Energyproduction cross-sections in e+e− annihilation at (1.45÷1.80) GeV c.m. Energy

B. Esposito, A. Marini, G. Piano-Mortari, F. Ronga, M. Nigro, L. Pescara, R. Bernabei, S. d'Angelo, P. Monacelli, M. Moricca, L. Paoluzi, R. Santonico, F. Sebastiani

Lettere Al Nuovo Cimento Series 2 · 1981

Vollständiger Abstract

Worum geht es in dieser Arbeit?

<h4>Objective</h4>This study investigated physicians' use of large language model (LLM)-generated assessments in psychiatric symptom evaluation and factors influencing their acceptance of artificial intelligence (AI) suggestions.<h4>Methods</h4>Twelve psychiatrists evaluated 50 anonymized psychiatric case vignettes, rating the presence and severity of 73 symptoms using a 4-point Likert scale. After initial assessment, they reviewed GPT-4o's symptom ratings and could revise their evaluations. Accuracy was measured against a gold standard established by expert consensus. We computed initial (I), LLM (L), and revised (R) accuracy scores using percent agreement and weighted kappa. Performance improvement was measured by relative kappa increase. Regression analyses examined the influence of task difficulty and user competence on accuracy improvement.<h4>Results</h4>LLM assessments outperformed initial physician ratings in both agreement (90.4% vs. 82.9%, p<0.001) and kappa (0.713 vs. 0.559, p<0.001). After revision, physician performance improved (κ=0.651) but remained below LLM levels. The switch rate-cases where physicians revised in response to LLM disagreement-was modest (25.4%), indicating partial reliance on AI. Performance gains were positively associated with LLM accuracy and negatively associated with users' own baseline competence, suggesting that less confident users rely more on LLM assistance.<h4>Conclusion</h4>Physicians demonstrated limited but strategic trust in LLM outputs, adjusting their judgments more when the LLM was accurate or when their own competence was lower. Miscalibrated trust-excessive skepticism or overreliance-led to missed gains. Effective human-AI collaboration in psychiatry requires tools and training for accurate self-assessment and AI trust calibration to avoid the pitfalls of algorithm aversion or overreliance.

Abstract: PubMed · Datensatz

Bibliografischer Nachweis

Publikationsdaten

Autor:innen
B. Esposito, A. Marini, G. Piano-Mortari, F. Ronga, M. Nigro, L. Pescara, R. Bernabei, S. d'Angelo, P. Monacelli, M. Moricca, L. Paoluzi, R. Santonico, F. Sebastiani
Quelle
Lettere Al Nuovo Cimento Series 2
Publikation
1981-01-01
Band / Ausgabe
Nicht angegeben
Seiten
Nicht angegeben
ISSN / ISBN
1827-613X
Zitationen
9 laut Crossref
Referenzen
0 hinterlegt

Zitieren

Zitierfähiger Nachweis

B. Esposito, A. Marini, G. Piano-Mortari, F. Ronga, M. Nigro, L. Pescara, R. Bernabei, S. d'Angelo, P. Monacelli, M. Moricca, L. Paoluzi, R. Santonico, F. Sebastiani (1981). Measurement on $$\pi ^ + \pi ^ - \pi ^0 \pi ^0 $$ , $$\pi ^ + \pi ^ - \pi ^ + \pi ^ - \pi ^0 $$ , $$\pi ^ + \pi ^ - \pi ^ + \pi ^ - \pi ^0 \pi ^0 $$ , $$\pi ^ + \pi ^ - \pi ^ + \pi ^ - \pi ^ + \pi ^ - $$ production cross-sections in e+e− annihilation at (1.45÷1.80) GeV c.m. Energyproduction cross-sections in e+e− annihilation at (1.45÷1.80) GeV c.m. Energy. Lettere Al Nuovo Cimento Series 2. https://doi.org/10.30773/pi.2025.0330
RIS BibTeX CSL-JSON

Kontext

Themen, Förderung und Nutzung

Lizenzhinweise: Lizenz 1