HealthContested · G 69 / P 68
Source article: Systematic Bias in Comparative Evaluations of Machine Learning Versus Logistic Regression for Clinical Prediction Models: A Meta-Research Analysis Using Trauma Mortality as an Empirical Case
Problem
Apparent superiority of ML over logistic regression for trauma mortality prediction may be inflated by convergent practices including comparing best-of-several ML models to a single LR comparator, reliance on internal validation, selective reporting, and AUC-only synthesis.
Journal of Clinical EpidemiologyGain
Across 17 studies totaling 243,324 trauma patients, the best-performing ML model showed a small pooled AUC advantage over logistic regression for mortality prediction.
Journal of Clinical EpidemiologyHealthContested · G 74 / P 72
Source article: Predicting Conversion from Mild Cognitive Impairment to Alzheimer's Disease: A Systematic Review of Deep Learning Models for Early-Stage Disease Classification
Problem
Deep learning models for MCI-to-AD conversion face substantial barriers to routine clinical use due to heavy reliance on ADNI, lack of diverse multicenter data, overfitting, and poor interpretability.
Ageing Research ReviewsGain
Deep learning models showed promising and often strong performance for predicting conversion from mild cognitive impairment to Alzheimer's disease, supporting early diagnosis and timely therapeutic intervention.
Ageing Research ReviewsPolicyContested · G 60 / P 59
Source article: Gavin Newsom imposes strict new rules on AI, social media and chatbots for children
Problem
Critics argue AB 1709 functions as a ban for under-16s that cuts young people off from digital lifelines and speech without making them safer or healthier.
The GuardianGain
California's new package creates the strongest U.S. companion chatbot rules and bans addictive features like infinite scroll for under-16s to protect young users from harm.
The GuardianHealthContested · G 67 / P 64
Source article: The impact of digital technology, social media, and artificial intelligence on cognitive functions: a review
Problem
Digital devices, social media and AI tools influence brain function and cognitive abilities, with potential negative impacts on attention, memory and related functions.
Frontiers in CognitionGain
Digital devices, social media and AI tools have brought convenience and connectivity and can have positive impacts on cognitive functions including attention and memory.
Frontiers in CognitionHealthContested · G 68 / P 70
Source article: Machine-learning prediction of urine-culture positivity in a multicentre test-ordered cohort: Model development and internal validation
Problem
Models were validated only at sample level without patient or centre grouping, and exploratory risk strata were not evaluated for clinical utility or safety, so they do not establish symptomatic UTI or safe antibiotic decisions.
BJUI CompassGain
In 2530 paired urinalysis-culture records from three university hospitals, gradient-boosting models estimated culture positivity after urinalysis, with CatBoost achieving test-set AUC 0.858 and high specificity at the reported threshold.
BJUI CompassHealthContested · G 69 / P 73
Source article: Machine Learning for Mortality Prediction in Infective Endocarditis: A Systematic Review and Meta-Analysis
Problem
Half of included studies had identified risk of bias and clinical adoption remains limited, requiring multicenter prospective validation and interpretable frameworks before bedside use.
Cardiology in ReviewGain
Supervised ML models, especially ensemble methods, predicted all-cause mortality in adult infective endocarditis with pooled AUC 0.85 for both in-hospital/early and 6-month mortality, outperforming conventional scores.
Cardiology in ReviewHealthNegative state · G 67 / P 76
Source article: Artificial intelligence for lung disease quantification in systemic sclerosis-associated interstitial lung disease and other connective tissue disease-associated interstitial lung disease
Problem
Visual HRCT scoring remains reader-dependent and AI outputs lack prospective multicenter validation and protocol harmonization needed to serve as treatment-triggering biomarkers.
Current Opinion in RheumatologyGain
In systemic sclerosis-associated ILD, AI-based HRCT quantification stratifies FVC decline and long-term survival and correlates with lung function measures to predict mortality.
Current Opinion in RheumatologyHealthContested · G 70 / P 70
Source article: How Well Do AI Chatbots Understand Abnormal Anatomy: A Comparative Study Using Congenital Anomalies and Tumor Cases
Problem
Chatbots sometimes confused similar congenital anomalies and provided less detailed anatomical descriptions in complex tumor cases, requiring caution and verification by qualified professionals before clinical use.
Clinical AnatomyGain
In a 20-case test of congenital anomalies and tumors, ChatGPT, Gemini and Copilot achieved 80-95% diagnostic accuracy with detailed anatomical descriptions, suggesting potential as supplementary radiological diagnostic support.
Clinical AnatomyHealthContested · G 70 / P 69
Source article: Stakeholder perspectives on artificial intelligence in schizophrenia care
Problem
Participants identified trust as the central barrier to using an AI companion, driven by privacy concerns and vulnerabilities specific to schizophrenia.
Psychological MedicineGain
Participants with schizophrenia recognized an AI companion as a potentially accessible source of support between clinical visits.
Psychological MedicineHealthContested · G 71 / P 71
Source article: Effect of Large Language Model-Powered Virtual Standardized Patients on History-Taking Among Undergraduate Medical Students: Propensity-Matched Cohort Study
Problem
Students with medium and low baseline history-taking proficiency showed relatively limited score improvements from LLM-VSP self-practice, with practice frequency alone not independently predicting final performance.
JMIR Medical EducationGain
Undergraduate medical students who used LLM-powered virtual standardized patients as extracurricular self-practice achieved higher end-of-term history-taking performance at an OSCE with real standardized patients compared to routine instruction.
JMIR Medical EducationHealthContested · G 70 / P 73
Source article: Beyond the Algorithm: A Stewardship Framework for the Hand Surgeon Adopting Artificial Intelligence
Problem
Most hand surgery AI tools are deployed in unaudited workflows and rarely remeasured after release after testing only on training-like data, leaving the hand surgeon accountable for patient outcomes shaped by opaque models.
The Journal of Hand SurgeryGain
AI tools are entering hand surgery practice to read scaphoid and distal radius radiographs and to predict outcomes after carpal tunnel release.
The Journal of Hand SurgeryHealthContested · G 69 / P 73
Source article: A Quality Assessment Rubric for Artificial Intelligence-Generated Patient-Friendly Radiology Reports
Problem
AI tools translating radiology reports into plain language can produce translation errors that compromise comprehension and safety, causing reports to be graded unsafe and warrant withholding from patients.
American Journal of RoentgenologyGain
A five-attribute rubric for AI-generated patient-friendly radiology reports showed almost-perfect agreement between lay and radiologist team members and may provide a standardized safeguard before patient distribution.
American Journal of RoentgenologyEducationContested · G 71 / P 71
Source article: Efficiency vs. safety in AI-enabled medical education: an ethical analysis of AI as a bridge or a wedge
HealthContested · G 71 / P 72
Source article: Machine Learning for Autism Spectrum Disorder Prediction: A Review of Data Augmentation and Feature Selection Techniques
Problem
Machine learning models for autism spectrum disorder prediction that use data augmentation and feature selection have limited external validation and inadequate evaluation frameworks, reducing confidence in reported performance improvements and model generalizability.
Health Care ScienceGain
Data augmentation and feature selection techniques may improve robustness, predictive performance, and interpretability of machine learning models for autism spectrum disorder prediction and help address dataset scarcity.
Health Care ScienceEducationContested · G 70 / P 71
Source article: Developing validity arguments for artificial intelligence-based assessment: Balancing affordances and threats
Problem
AI-based assessment introduces distinct validity threats across scoring, generalisation, extrapolation and implications, including contamination, instability, inequities, automation bias and deskilling when used for consequential learner progression decisions.
Medical EducationGain
AI systems can generate, score and interpret educational assessments that inform learner progression, with design and governance determining whether cross-cutting mechanisms function as affordances.
Medical Education