ChatGPT-4o answered expert-validated true/false questions on vertical root fractures and tooth cracks with 86.1% accuracy, outperforming ChatGPT-3.5 and Gemini in a 5,400-response repeated test.
A peer-reviewed study in Odontology compared ChatGPT-3.5, ChatGPT-4o, and Google Gemini on 60 expert-validated true/false questions about longitudinal tooth fractures, querying each model three times daily for 10 days for 5,400 total responses against expert reference answers.
- Impact 30%
- 69
- Evidence 25%
- 95
- Scale 20%
- 35
- Confidence 15%
- 87
- Recency 10%
- 99
Updated Oct 7, 2026 · TRV-2026-1305
ace