EUVIMEDEuropean Health Evidence
Uhr 7/7Sources Journal Tree
Easy Demo

Lokaler Crossref-Datenbestand · journal-article

Large Language Models Versus Multidisciplinary Tumor Board Decisions in Thyroid Cancer

Bayram Barış Büyük, Arzu Or Koca, Felat Toprak, Aslı Ayan, Tuncay Tunçcan, Ayça Arslan, ArzuBetül Duran, Elif Akyol Şen, Ömer Semih Budak, Alptekin Gürsoy, Caner Kılıç

Laryngoscope Investigative Otolaryngology · 2026

Vollständiger Abstract

Worum geht es in dieser Arbeit?

ABSTRACT Objectives Large language models (LLMs) are increasingly proposed as clinical decision‐support tools; however, their agreement with real‐world multidisciplinary tumor board (MDT) decisions remains insufficiently investigated in thyroid oncology. To evaluate the concordance between treatment recommendations generated by ChatGPT 5.2 and Gemini 3.0 and decisions made by a tertiary multidisciplinary thyroid tumor board. Methods This study included 59 consecutive patients discussed at a tertiary MDT between August and December 2025. Anonymized clinical data, including demographics, ultrasonographic findings, and Bethesda cytology, were provided to both LLMs using standardized structured prompts. MDT decisions were defined as the reference standard. Agreement was assessed using exact concordance rates and Cohen's kappa ( κ ) statistics with 95% confidence intervals. Results ChatGPT 5.2 achieved a concordance rate of 71.2% (42/59), demonstrating substantial agreement ( κ = 0.623; 95% CI 0.459–0.771). Gemini 3.0 showed a concordance rate of 64.4% (38/59), reflecting moderate agreement ( κ = 0.527; 95% CI 0.359–0.684). Discordance increased in complex scenarios involving lateral neck dissection, radioactive iodine therapy, and active surveillance. Conclusions While LLMs demonstrate promising concordance in standardized thyroid cancer management, they are best positioned as supportive decision aids—such as in MDT preparation and workflow streamlining—rather than replacements for expert multidisciplinary evaluation, particularly in complex clinical scenarios. Level of Evidence 3.

Bibliografischer Nachweis

Publikationsdaten

Autor:innen
Bayram Barış Büyük, Arzu Or Koca, Felat Toprak, Aslı Ayan, Tuncay Tunçcan, Ayça Arslan, ArzuBetül Duran, Elif Akyol Şen, Ömer Semih Budak, Alptekin Gürsoy, Caner Kılıç
Quelle
Laryngoscope Investigative Otolaryngology
Publikation
2026-01-01
Band / Ausgabe
Nicht angegeben
Seiten
Nicht angegeben
ISSN / ISBN
2378-8038, 2378-8038
Zitationen
0 laut Crossref
Referenzen
0 hinterlegt

Zitieren

Zitierfähiger Nachweis

Bayram Barış Büyük, Arzu Or Koca, Felat Toprak, Aslı Ayan, Tuncay Tunçcan, Ayça Arslan, ArzuBetül Duran, Elif Akyol Şen, Ömer Semih Budak, Alptekin Gürsoy, Caner Kılıç (2026). Large Language Models Versus Multidisciplinary Tumor Board Decisions in Thyroid Cancer. Laryngoscope Investigative Otolaryngology. https://doi.org/10.1002/lio2.70547
RIS BibTeX CSL-JSON

Kontext

Themen, Förderung und Nutzung

Lizenzhinweise: Lizenz 1 · Lizenz 2