TruaceTracing the truth around AITuesday, September 15, 2026
The Index

What the evidence says.What the public feels.

Ranks distinct AI gain and problem claims from the published record. Scores reward impact, independent source strength, scale, confidence, and recency.

1,419 results
Show filters and sorting

AI gains · 788

67
GainHealth· Stable· Evidence: Moderate (1 source)

Deep learning models, especially U-Net variants and newer transformer ensembles, improved ischemic stroke lesion segmentation on MRI, achieving Dice scores above 0.80 and approaching 0.90 in multisite DWI reports, to support faster treatment selection and quantification.

By July 2026, a narrative review of 40 studies from 2020-2025 found deep learning, led by U-Net variants with residual and attention mechanisms and standardized pipelines like nnU-Net, increasingly achieved high Dice scores on MRI DWI/ADC, with many reports above 0.80 and recent transformer and ensemble multisite models approaching 0.90, while CT performance was lower and more variable.

Impact 30%49
Evidence 25%95
Scale 20%35
Confidence 15%87
Recency 10%89

Updated Jul 19, 2026 · TRV-2026-0268

67
GainHealth· Stable· Evidence: Moderate (1 source)

Explainable ML models used after Wells/Geneva triage improved early PTE prediction in ED patients, with Extra Trees reaching 0.82 accuracy and 0.83 AUC.

In a 2022-2024 study of 472 emergency department patients with suspected pulmonary thromboembolism across three Mashhad University of Medical Sciences centers, researchers developed explainable machine-learning models intended to operate after Wells/Geneva triage. Using CTPA as reference, Extra Trees achieved accuracy 0.82, sensitivity 0.69, specificity 0.86 and AUC 0.83 for PTE prediction, and AUC 0.77 for central and 0.67 for peripheral emboli for anatomical classification.

Impact 30%49
Evidence 25%95
Scale 20%35
Confidence 15%87
Recency 10%89

Updated Jul 19, 2026 · TRV-2026-0267

67
GainHealth· Stable· Evidence: Moderate (1 source)

LLM chatbots answered periodontal patient questions with scientific accuracy comparable to expert periodontologists while scoring higher on completeness and empathy.

A July 2026 peer-reviewed study compared ChatGPT GPT-5.1, Gemini 2.5 Flash, and Claude Sonnet 4.5 against expert periodontologists on 20 periodontal patient questions. Nine blinded periodontologists rated anonymized answers for scientific accuracy, completeness, conciseness & focus, empathy, and clarity.

Impact 30%49
Evidence 25%95
Scale 20%35
Confidence 15%87
Recency 10%89

Updated Jul 19, 2026 · TRV-2026-0265

67
GainHealth· Stable· Evidence: Moderate (1 source)

LightGBM model trained on 282 Lenke 1/2 AIS patients predicted postoperative coronal imbalance with AUC 0.885 training and 0.824 internal validation, identifying LIV-LSTV, Lumbar Modifier, and Risser grade as key predictors.

In 282 patients with Lenke 1/2 adolescent idiopathic scoliosis treated with selective posterior thoracic fusion, investigators built an interpretable machine learning pipeline to stratify risk of postoperative coronal imbalance. After reducing 24 candidates to key features, a LightGBM model achieved the highest discrimination with AUC 0.885 in training and 0.824 in internal validation, with SHAP highlighting LIV-LSTV relationship, Lumbar Modifier, and Risser grade.

Impact 30%49
Evidence 25%95
Scale 20%35
Confidence 15%87
Recency 10%89

Updated Jul 19, 2026 · TRV-2026-0263

AI problems · 631

55
ProblemEducation· Stable· Evidence: Moderate (1 source)

Assigning sixth-grade students in a Brooklyn middle school to use Google Gemini for feedback on a science experiment risks teaching students to outsource thinking to machines instead of peer discussion and revision

In October, a sixth-grade student at a middle school in Brooklyn received an assignment to create a science experiment and then ask Google Gemini for feedback. His mother, Kelly Clancy, objected to the teacher and later founded Parents for AI Caution in Educational Spaces, which is pushing for a two-year moratorium on AI in New York City public schools.

Impact 30%49
Evidence 25%62
Scale 20%35
Confidence 15%62
Recency 10%88

Updated Jul 12, 2026 · TRV-2026-0087

55
ProblemCrime· Stable· Evidence: Moderate (1 source)

Criminals using artificial intelligence to run investment scams persuaded UK victims to move money into fake funds, causing about £221.5m in losses in 2025

UK Finance reported almost 15,000 investment scams in 2025, with about £221.5m lost after people were persuaded to move money to fake investments or fictitious funds, a 40% rise on the prior year. The report said criminals use artificial intelligence to dupe people and that advances in AI make large-scale scams easier, with schemes involving gold, cryptocurrencies and wine.

Impact 30%49
Evidence 25%62
Scale 20%35
Confidence 15%62
Recency 10%88

Updated Jul 12, 2026 · TRV-2026-0084

55
ProblemSports· Stable· Evidence: Moderate (1 source)

At Wimbledon 2025, the new AI electronic line-judging system failed to spot an out ball hit long by Sonay Kartal.

At Wimbledon in July 2025, organizers replaced 300 human line judges with an artificial intelligence electronic line-judging system. Shortly after deployment, the new system failed to detect that player Sonay Kartal had hit a ball long during a match.

Impact 30%49
Evidence 25%62
Scale 20%35
Confidence 15%62
Recency 10%88

Updated Jul 12, 2026 · TRV-2026-0083

55
ProblemCrime· Stable· Evidence: Moderate (1 source)

Criminals exploiting AI technology to take over mobile, banking and online shopping accounts contributed to a record 444,000 fraud reports to the UK national fraud database last year.

Cifas, the UK's fraud prevention organisation, reported 444,000 fraud cases from its members last year, a 6% rise on 2024 and a record for its national fraud database. The body said criminals are increasingly exploiting AI technology to take over mobile, banking and online shopping accounts, using stolen data to make unauthorised transactions and enabling deception on industrialised levels.

Impact 30%49
Evidence 25%62
Scale 20%35
Confidence 15%62
Recency 10%88

Updated Jul 12, 2026 · TRV-2026-0081

Recomputed live from the record · Sep 15, 2026, 6:14 PM