LaborContested · G 72 / P 74
Source article: Emotional AI and the rise of pseudo-intimacy: are we trading authenticity for algorithmic affection?
Problem
Emotional AI companions risk replacing authentic human intimacy with algorithmic pseudo-intimacy, fostering emotional dependence, sustained loneliness, and displacement of human ties.
Frontiers in PsychologyGain
Emotional AI systems like Replika and Xiaoice offer accessible companionship and promise of connection for users facing isolation or psychological distress.
Frontiers in PsychologyEducationContested · G 70 / P 69
Source article: Exploring the Impact of Generative AI ChatGPT on Critical Thinking in Higher Education: Passive AI-Directed Use or Human–AI Supported Collaboration?
Problem
When used in passive AI-directed mode, ChatGPT left students neutral at the resolution stage of cognitive presence, creating a resolution gap in critical thinking development.
Education SciencesGain
When used with guidance in collaborative human-AI interaction, ChatGPT helped university students complete all four cognitive presence phases and enhanced critical thinking, with self-reported improvements in triggering, exploration, and integration.
Education SciencesPolicyContested · G 55 / P 54
Source article: The UK’s G20 presidency will push Andy Burnham to the centre of the global stage. What will he use it for? | Michael Jacobs
Problem
Recent alarm over existential risks of AI has created pressure for international control, but no global regulatory agreement exists.
The GuardianGain
The UK's upcoming G20 presidency could be used to secure agreement on a process toward a global treaty regulating artificial intelligence.
The GuardianHealthContested · G 65 / P 68
Source article: Consistency and accuracy of different artificial intelligence models in evaluating longitudinal cracks and fractures of teeth
Problem
The same models showed very low inter-model agreement and significant temporal instability, with Gemini performing worse in evening sessions, indicating they cannot replace clinician judgment for vertical root fracture decisions.
OdontologyGain
ChatGPT-4o answered expert-validated true/false questions on vertical root fractures and tooth cracks with 86.1% accuracy, outperforming ChatGPT-3.5 and Gemini in a 5,400-response repeated test.
OdontologyClimateNegative state · G 51 / P 56
Source article: Making scents: could AI help perfumers take pressure off endangered plants?
Problem
Recreating endangered plant scents with AI and biotechnology risks greenwashing because synthetic versions do not save the species or eliminate demand for wild-harvested material.
The GuardianGain
AI trained on millions of human smell responses can predict and reconstruct fragrances from tiny samples, enabling lab production of compounds from endangered plants without repeated harvesting.
The GuardianPolicyNegative state · G 68 / P 73
Source article: Generative AI and misinformation: a scoping review of the role of generative AI in the generation, detection, mitigation, and impact of misinformation
Problem
LLMs can generate highly convincing misinformation that exploits audience biases, and exposure was found to reduce trust and influence decision-making.
AI & SOCIETYGain
LLMs demonstrated capacity to detect false claims and increase users' resistance to misinformation, with personalized corrections showing effectiveness.
AI & SOCIETYHealthNegative state · G 65 / P 70
Source article: Engineering smart polymeric lipid nanoparticles for breast cancer: The convergence of AI and medical personalization to enhance efficacy and safety
Problem
No AI-designed polymeric-lipid nanoparticle formulation for breast cancer has entered registered clinical trials, leaving scalable production and regulatory clearance unresolved.
International Journal of PharmaceuticsGain
Preclinical studies indicate AI-optimized, ligand-functionalized polymeric-lipid nanoparticles can improve tumor-specific accumulation and reduce systemic side effects for breast cancer subtypes.
International Journal of PharmaceuticsHealthPositive state · G 80 / P 75
Source article: Unfolding new horizons: Machine learning applications for pediatric intussusception - A systematic review and meta-analysis
Problem
Review detected possible publication bias, found no studies reporting clinical endpoint data, and found abdominal radiograph triage models had substantially lower external accuracy, limiting readiness for clinical implementation.
Journal of Pediatric SurgeryGain
Systematic review of 11 studies (36,863 patients) found ultrasound-based ML models achieved high pooled sensitivity and specificity with external validation, and AI assistance improved junior readers' specificity and cut examination time by 61% without loss of accuracy.
Journal of Pediatric SurgeryHealthContested · G 71 / P 69
Source article: Guideline-augmented prompting improves comparative preference and response consistency of large language model outputs for orthopaedic anaesthesia questions: A controlled prompting study
Problem
LLM responses for orthopaedic anaesthesia showed variable consistency, with later models exhibiting greater partial variability despite avoiding fully contradictory outputs.
Journal of Experimental OrthopaedicsGain
Guideline-augmented prompting increased blinded preference win rates and improved response consistency for LLMs answering orthopaedic anaesthesia questions compared to human experts.
Journal of Experimental OrthopaedicsHealthContested · G 67 / P 70
Source article: Development and Nationwide Multicentre Evaluation of Guideline-Grounded Large Language Model Chatbots to Support Patient Self-Management and Education in Rheumatology
Problem
Chatbots failed to answer 4.2% of interactions, received negative ratings for insufficient detail, and showed only 45% full guideline adherence with weak agreement between LLM and physician assessments.
Journal of Medical SystemsGain
Patients using guideline-grounded rheumatology chatbots reported high usability and satisfaction, with most answers rated safe and correct in real-world deployment.
Journal of Medical SystemsMedia & ArtsContested · G 54 / P 57
Source article: Sony reportedly demanded music platforms take down over 260,000 AI deepfake tracks.
Problem
AI music platforms hosted AI tracks imitating Sony artists, leading to copyright infringement claims.
The VergeGain
Sony Music nearly doubled its enforcement volume against AI-generated impersonations of its artists.
The VergeHealthContested · G 71 / P 69
Source article: [Atrial fibrillation screening-reading the P-wave]
Problem
An AI-based P-wave risk score derived from sinus-rhythm ECG does not establish an AF diagnosis and alone does not constitute an indication for treatment, risking overinterpretation in screening.
Die Innere MedizinGain
AI models analyzing sinus-rhythm ECGs can identify atrial risk phenotypes and help select patients for prolonged rhythm monitoring to improve AF detection yield.
Die Innere MedizinHealthContested · G 71 / P 73
Source article: Generative AI versus physicians in diagnostic radiology: a systematic review and meta-analysis
Problem
Generative AI showed significantly lower diagnostic accuracy than expert physicians on radiology tasks, with a 13.0 percentage point gap.
Japanese Journal of RadiologyGain
Generative AI achieved higher diagnostic accuracy when given text-only input compared to image-only input in radiology tasks.
Japanese Journal of RadiologyHealthContested · G 69 / P 69
Source article: Explainable machine learning for identifying mild cognitive impairment in older adults with chronic diseases: Model development and temporal validation
Problem
Model had limited sensitivity and calibration, with sensitivity dropping to 0.390 in temporal validation and Brier score worsening to 0.137, limiting clinical utility.
Acta PsychologicaGain
XGBoost model achieved moderate and relatively stable discrimination for identifying current MCI status among older adults with chronic conditions across 2018 internal and 2020 temporal validation, with highly stable TreeSHAP explanations.
Acta PsychologicaHealthNegative state · G 66 / P 71
Source article: Rule-Based Versus Generative Extraction of Psychological Symptoms From Forensic Medical Certificates: A Feasibility Study in the French ORFeAD Network
Problem
Sensitivity was heterogeneous and low for low-prevalence symptoms, 11 of 35 variables failed the reliability threshold under either pipeline, and neither pipeline supports individual-level decisions, with the generative model incurring far higher compute cost.
Behavioral Sciences & the LawGain
Automated extraction of psychological and subjective variables from unstructured forensic certificates was feasible with high specificity, with 24 of 35 variables meeting an 85% reliability threshold under at least one pipeline.
Behavioral Sciences & the Law