In blinded evaluation of 24 urogynecology patient queries, ChatGPT Health achieved higher overall patient-facing quality scores than ChatGPT Plus, with significantly greater clarity, completeness, and usefulness, while maintaining 100% median guideline concordance.
A blinded comparison published 2 October 2026 tested ChatGPT Health against standard ChatGPT Plus on 24 guideline-informed urogynecology patient queries. Five clinicians anonymously rated 48 responses for quality, guideline concordance, and safety, with paired Wilcoxon tests at the prompt level.
- Impact 30%
- 63
- Evidence 25%
- 95
- Scale 20%
- 35
- Confidence 15%
- 87
- Recency 10%
- 99
Updated Oct 5, 2026 · TRV-2026-1287
ace