Concordance between multidisciplinary tumor board decisions and AI-based recommendations in endometrial cancer: impact of discordance direction on clinical outcomes.
ABSTRACT Objectives Large language models (LLMs) are increasingly proposed as clinical decision‐support tools; however, their agreement with real‐world multidisciplinary tumor board (MDT) decisions remains insufficiently investigated in thyroid oncology. To evaluate the concordance between treatment recommendations gen...
B. B. Büyük, Arzu Or Koca, F. Toprak et al.· Laryngoscope Investigative O...· 0 citations
Large language models (LLMs) such as GPT-4 are being evaluated for their use as supportive tools in oncological treatment planning. However, in pancreatic cancer, current studies are confined to predefined question–answer formats, while studies specifically investigating real-world scenarios that benchmark LLM performa...
F. Gehrisch, K. Kirkgöz, Antonie Willner et al.· Langenbeck's archives of sur...· 0 citations
Both models demonstrated high agreement with expert GIST MTB recommendations, with no significant performance difference between them, and support a potential assistive role for LLMs in GIST MTB workflows, while underscores the continued necessity of expert oversight.
PURPOSE
Molecular classification has refined risk stratification in endometrial cancer and is now incorporated into the 2023 International Federation of Gynecology and Obstetrics (FIGO) staging system. However, prospective real-world data on its impact on multidisciplinary tumor board (MDT) decision making remain limit...
R. Pinninti, H. Abbaraju, Krishna Mohan Mallavarapu et al.· JCO Global Oncology· 1 citation
PURPOSE
To compare the concordance of ChatGPT, Gemini, and Claude with a prespecified expert guideline-based reference standard in fabricated endometrial cancer clinical vignettes under standardized prompting.
METHODS
We conducted a case-based in silico benchmarking study using 35 fabricated postoperative endometrial...
E. Perrone, Giuseppe Parisi, M. Giuliano et al.· JCO Clinical Cancer Informat...· 0 citations
ChatGPT achieved the highest overall concordance, although all models generated clinically acceptable recommendations in most cases, and all LLMs demonstrated high concordance with consultant-led thyroid cancer MDT decisions.
A. White, Kerry A. Leyton, N. Patel et al.· Updates in Surgery· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.