Ranks distinct AI gain and problem claims from the published record. Scores reward impact, independent source strength, scale, confidence, and recency.
Researchers retrospectively tested ChatGPT on 300 histopathologically confirmed oral lichen planus cases with at least 24 months of follow-up, using serial clinical records, intraoral photographs, and histopathology reports. Compared with blinded expert panel consensus, the model achieved 94.7% accuracy for trajectory classification and 78.8% sensitivity with 99.6% specificity for high-risk detection as of the August 2026 publication.
A comparative study merged SEE-AI and Kvasir-Capsule into a 21-class capsule endoscopy image dataset and fine-tuned a Vision Transformer, DenseNet121, and ResNet50. On an independent test set of 8,696 frames, the transformer achieved 92.2% accuracy and 0.99 AUC, substantially higher than the two CNN baselines under the reported experimental conditions.
A retrospective study of 2293 ADNI brain MRIs plus 270 external NACC scans tested whether SubtleHD, an FDA-cleared deep learning enhancement tool, could improve downstream Alzheimer's classification when applied to already diagnostic-quality 1.5T T1-weighted images. ResNet34 and DenseNet121 models trained on enhanced images outperformed those trained on standard-of-care images on internal and external tests.
On 2026-08-07, researchers reported a urine-based test for urothelial bladder cancer that combines solvent extraction, GC-MS profiling, and machine learning. In 100 participants, an XGBoost model using an 8-metabolite panel achieved AUROC 0.869, improving on classical statistics at 0.752, with 85% balanced sensitivity and specificity.
By July 2026, the authors described an AI-investment paradox in enterprises: continued heavy investment alongside initiatives that fail to scale. They proposed a decision-centric portfolio framework that reframes governance around discrete investable decision opportunities within workflows, introducing AI-Investable Process Nodes as bounded points where benefits, risks and costs can be assessed ex ante.
Researchers analyzed 593 tasks across 126 occupations in the cultural and creative industries using GPT-4 generated synthetic annotations of Australian Skills Classification descriptions. They measured cognitive and behavioural rules and estimated AI autonomy feasibility and efficiency potential to map where human, AI, or hybrid carriers fit.
In a retrospective study of 102 patients at a tertiary care center in India, researchers used k-means clustering on eight biopsychosocial baseline variables to derive three AUD profiles. By the August 2026 publication date, they reported Late-Onset, High-Functioning, and Severe groups with differing 3-month abstinence rates corroborated by GGT levels and bootstrap-assessed cluster stability.