TruaceTracing the truth around AIWednesday, July 22, 2026
TRV-2026-0412Version 1 · Certified

Written 2026-07-20 10:35:39 UTC · current record

Reason for this version

Certified into the record

Canonical text (the exact bytes fingerprinted)

TRUVACE RECORD VERSION
record: TRV-2026-0412
version: 1
kind: certified
reason: Certified into the record
timestamp: 2026-07-20T10:35:39.388684Z
status: published
lens: trace
sector: health
headline: ChatGPT Provides Satisfactory but Occasionally Inaccurate Answers to Common Patient Hip Arthroscopy Questions
dek: PURPOSE: To assess the ability of ChatGPT to answer common patient questions regarding hip arthroscopy, and to analyze the accuracy and appropriateness of its responses. METHODS: Ten questions were selected from well-known patient education websites, and ChatGPT (version 3.5) responses to these questions were graded by 2 fellowship-trained hip preservation surgeons. Responses were analyzed, compared with the current literature, and graded from A to D (A being the highest, and D being the lowest) in a grading sca…
gain_title: ChatGPT provided satisfactory answers to common patient hip arthroscopy questions, with half of responses graded A and another 30% graded B by fellowship-trained surgeons.
problem_title: ChatGPT responses contained incorrect information in more than one instance and were written at a college graduate reading level, requiring caution for patient education.
trace_subject: accuracy of ChatGPT answers to common patient hip arthroscopy questions
gain_reading: ChatGPT provided satisfactory answers to common patient hip arthroscopy questions, with half of responses graded A and another 30% graded B by fellowship-trained surgeons.
gain_evidence: ChatGPT can answer common patient questions regarding hip arthroscopy with satisfactory accuracy graded by 2 high-volume hip arthroscopists
problem_reading: ChatGPT responses contained incorrect information in more than one instance and were written at a college graduate reading level, requiring caution for patient education.
problem_evidence: incorrect information was identified in more than one instance | Caution must be observed when using ChatGPT for patient education related to hip arthroscopy
quick_read: In a study published June 22, 2024, two hip preservation surgeons graded ChatGPT 3.5 answers to ten common hip arthroscopy questions drawn from patient education sites, using an A-to-D scale and readability scores FRES and FKGL.

The findings matter because patients increasingly turn to chatbots for surgical information, yet the mix of mostly satisfactory grades alongside documented inaccuracies and college-level readability raises questions about safe use without clinician oversight.
limitation: 
tag: Automated dual reading
key_points: Ten common patient questions from education websites were answered by ChatGPT version 3.5 and graded A to D by two fellowship-trained hip preservation surgeons. | Consensus grades were A 50%, B 30%, C 10%, D 10%, with initial inter-rater agreement of 30%. | Readability analysis found mean Flesch-Kincaid Reading Ease Score 28.2 and mean Grade Level 14.4, indicating college-level text.
rundown: Researchers selected ten common hip arthroscopy questions from patient education websites and prompted ChatGPT 3.5, then had two fellowship-trained hip preservation surgeons grade answers A to D for accuracy and completeness, reaching consensus when needed.

Results showed five As, three Bs, one C and one D, with mean Flesch-Kincaid Reading Ease 28.2 and Grade Level 14.4, and authors noted potential to aid physicians while warning about inaccuracies.
sources:
- peer_reviewed | Arthroscopy | https://doi.org/10.1016/j.arthro.2024.06.017 | 2024-06-22
prev: 0000000000000000000000000000000000000000000000000000000000000000
sha256
24d86aa02c5b355e773f597d21591968f968e48b4e665358c535096d3fa2150e
previous
0000000000000000000000000000000000000000000000000000000000000000
Verify this record
How to verify without trusting this page

Fetch the canonical text of any version from /api/record/TRV-2026-0412 and hash it yourself — for example shasum -a 256 on the saved canonical field. The result must equal content_hash, and each version’s text ends with prev:followed by the prior version’s hash (version 1 chains to 64 zeros). If a single character of any version had been altered since certification, the chain would not reproduce.