TruaceTracing the truth around AITuesday, July 21, 2026
TRV-2026-0321Version 1 · Certified

Written 2026-07-20 08:46:27 UTC · current record

Reason for this version

Certified into the record

Canonical text (the exact bytes fingerprinted)

TRUVACE RECORD VERSION
record: TRV-2026-0321
version: 1
kind: certified
reason: Certified into the record
timestamp: 2026-07-20T08:46:27.786407Z
status: published
lens: trace
sector: health
headline: Artificial Intelligence for Evidence Synthesis of Emerging Biologics to Improve Skeletal Health in Osteogenesis Imperfecta: Systematic Review and Meta-Analysis
dek: Background: Osteogenesis imperfecta (OI) is a rare genetic disorder characterized by bone fragility and recurrent fractures. Emerging biologics demonstrate promise by targeting bone-remodeling pathways, yet evidence for their efficacy and safety remains fragmented and heterogeneous, and no prior systematic review in OI has incorporated artificial intelligence (AI) to synthesize it. Objective: This study aims to systematically evaluate the efficacy and safety of novel biologics in patients with OI using an AI-ass…
gain_title: GPT-4o integrated into systematic review screening substantially accelerated evidence synthesis for osteogenesis imperfecta biologics while maintaining high sensitivity.
problem_title: GPT-4o showed optimism and positional biases in risk-of-bias assessment because it relied on probabilistic language patterns rather than structured clinical reasoning.
trace_subject: GPT-4o-assisted evidence synthesis for osteogenesis imperfecta biologics
gain_reading: GPT-4o integrated into systematic review screening substantially accelerated evidence synthesis for osteogenesis imperfecta biologics while maintaining high sensitivity.
gain_evidence: AI achieved high sensitivity in abstract (97.4%) and full-text (88.9%) screening
problem_reading: GPT-4o showed optimism and positional biases in risk-of-bias assessment because it relied on probabilistic language patterns rather than structured clinical reasoning.
problem_evidence: the model exhibited optimism and positional biases due to reliance on probabilistic language patterns rather than structured clinical reasoning
quick_read: By December 2025, researchers conducted a systematic review and meta-analysis of 13 trials (n=684) of five emerging biologics for osteogenesis imperfecta, using GPT-4o to perform parallel title/abstract and full-text screening and to assist risk-of-bias assessment. The AI workflow achieved 97.4% sensitivity at abstract level and 88.9% at full-text, reducing total screening time by over 95% with substantial agreement to humans (kappa 0.778).

The acceleration matters for rare-disease research where evidence is fragmented, but clinical impact remains limited: denosumab and setrusumab improved lumbar spine aBMD without demonstrating superior fracture reduction versus bisphosphonates, and safety signals like 30.95% hypercalcemia with denosumab in children persist. The observed optimism and positional biases indicate that scaling this workflow will still require explicit human oversight for contextual clinical reasoning.
limitation: Findings are constrained by a small and heterogeneous trial base, limiting generalizability of both efficacy estimates and AI performance.
tag: Model-prefilled trace
key_points: Systematic review included 13 trials (n=684) of denosumab, setrusumab, teriparatide, romosozumab, and fresolimumab up to December 1, 2025, with 10 trials (n=333) in meta-analysis. | In children denosumab produced 25.49% increase in lumbar spine aBMD at 12 months; in adults setrusumab yielded 9.38% improvement. | No biologic significantly reduced fracture incidence compared to bisphosphonates across trials. | Safety varied: denosumab associated with 30.95% hypercalcemia risk in children, while setrusumab had no treatment-related serious adverse events.
rundown: The review searched PubMed, Web of Science, Embase, ScienceDirect, Cochrane Library and ClinicalTrials.gov to December 1, 2025, including randomized, nonrandomized and single-arm trials reporting aBMD and/or fractures, excluding case series.

GPT-4o performed parallel 2-stage screening and assisted risk of bias assessment using an adapted Cochrane RoB 2 tool, benchmarked against humans with sensitivity, specificity and weighted Cohen kappa of 0.778.

Primary outcome was percentage change in aBMD synthesized via random-effects meta-analysis, showing age-specific gains but no superior fracture reduction over bisphosphonates.
sources:
- peer_reviewed | Journal of Medical Internet Research | https://doi.org/10.2196/85840 | 2026-07-10
prev: 0000000000000000000000000000000000000000000000000000000000000000
sha256
9efe60cc27b6979d0da0f3047e81e47403a3fcafe2f70532b09a58c8f906f43c
previous
0000000000000000000000000000000000000000000000000000000000000000
Verify this record
How to verify without trusting this page

Fetch the canonical text of any version from /api/record/TRV-2026-0321 and hash it yourself — for example shasum -a 256 on the saved canonical field. The result must equal content_hash, and each version’s text ends with prev:followed by the prior version’s hash (version 1 chains to 64 zeros). If a single character of any version had been altered since certification, the chain would not reproduce.