Evaluation of AI Chatbot Responses to Pediatric Urology Frequently Asked Questions.
OBJECTIVE To evaluate the quality of responses from four publicly available LLMs (ChatGPT-4o, Claude 3.7, Gemini 2.5, and Copilot) to frequently asked questions (FAQs) in pediatric urology. METHODS FAQs were generated using standardized prompts and submitted to each LLM using parent-centered instructions. Two board-c...