Open access
Jul 2026
Diagnostic capability of large language models in critically ill patients: a prospective single-centre study comparing ChatGPT, Claude, and Gemini with emergency physicians.
The LLMs achieved moderate agreement with ED reference diagnoses in critically ill patients but were consistently outperformed by physicians at the early diagnostic phases; their current diagnostic role in the ED remains limited.
İbrahim Günaydın, M. Yılmaz, Sinan Akpunar et al.
· BMC Emergency Medicine · 0 citations