Open access
Jul 2026
Evaluation and comparison of large language model responses to patient questions after diagnosis of high-risk human papillomavirus infection: an expert-rated digital patient education study
ChatGPT-5.5 Instant achieved higher ratings in several quality domains, but response-level binary safety differences were imprecise and not statistically significant, but both models require guideline-based clinical oversight.
Zhen Hao, Lin Wang, Yue Wu et al.
· Frontiers in Public Health · 0 citations