All traces

Download every Trace with both directional scores:Export CSVExport JSON

Emotional AI and the rise of pseudo-intimacy: are we trading authenticity for algorithmic affection?
LaborContested · G 72 / P 74

emotional AI companion systems providing connection to isolated users and their impact on authentic intimacy

Source article: Emotional AI and the rise of pseudo-intimacy: are we trading authenticity for algorithmic affection?

Problem

Emotional AI companions risk replacing authentic human intimacy with algorithmic pseudo-intimacy, fostering emotional dependence, sustained loneliness, and displacement of human ties.

Frontiers in Psychology
Gain

Emotional AI systems like Replika and Xiaoice offer accessible companionship and promise of connection for users facing isolation or psychological distress.

Frontiers in Psychology
Exploring the Impact of Generative AI ChatGPT on Critical Thinking in Higher Education: Passive AI-Directed Use or Human–AI Supported Collaboration?
EducationContested · G 70 / P 69

GenAI ChatGPT use by university students affecting critical thinking across cognitive presence phases

Source article: Exploring the Impact of Generative AI ChatGPT on Critical Thinking in Higher Education: Passive AI-Directed Use or Human–AI Supported Collaboration?

Problem

When used in passive AI-directed mode, ChatGPT left students neutral at the resolution stage of cognitive presence, creating a resolution gap in critical thinking development.

Education Sciences
Gain

When used with guidance in collaborative human-AI interaction, ChatGPT helped university students complete all four cognitive presence phases and enhanced critical thinking, with self-reported improvements in triggering, exploration, and integration.

Education Sciences
The UK’s G20 presidency will push Andy Burnham to the centre of the global stage. What will he use it for? | Michael Jacobs
PolicyContested · G 55 / P 54

global treaty process for regulation of artificial intelligence to address existential risks

Source article: The UK’s G20 presidency will push Andy Burnham to the centre of the global stage. What will he use it for? | Michael Jacobs

Problem

Recent alarm over existential risks of AI has created pressure for international control, but no global regulatory agreement exists.

The Guardian
Gain

The UK's upcoming G20 presidency could be used to secure agreement on a process toward a global treaty regulating artificial intelligence.

The Guardian
Consistency and accuracy of different artificial intelligence models in evaluating longitudinal cracks and fractures of teeth
HealthContested · G 65 / P 68

AI models answering dichotomous clinical questions on vertical root fractures and tooth cracks

Source article: Consistency and accuracy of different artificial intelligence models in evaluating longitudinal cracks and fractures of teeth

Problem

The same models showed very low inter-model agreement and significant temporal instability, with Gemini performing worse in evening sessions, indicating they cannot replace clinician judgment for vertical root fracture decisions.

Odontology
Gain

ChatGPT-4o answered expert-validated true/false questions on vertical root fractures and tooth cracks with 86.1% accuracy, outperforming ChatGPT-3.5 and Gemini in a 5,400-response repeated test.

Odontology
Making scents: could AI help perfumers take pressure off endangered plants?
ClimateNegative state · G 51 / P 56

using AI and biotechnology to reproduce scents from endangered fragrance plants to reduce wild harvesting

Source article: Making scents: could AI help perfumers take pressure off endangered plants?

Problem

Recreating endangered plant scents with AI and biotechnology risks greenwashing because synthetic versions do not save the species or eliminate demand for wild-harvested material.

The Guardian
Gain

AI trained on millions of human smell responses can predict and reconstruct fragrances from tiny samples, enabling lab production of compounds from endangered plants without repeated harvesting.

The Guardian
Generative AI and misinformation: a scoping review of the role of generative AI in the generation, detection, mitigation, and impact of misinformation
PolicyNegative state · G 68 / P 73

LLMs handling misinformation for general audiences

Source article: Generative AI and misinformation: a scoping review of the role of generative AI in the generation, detection, mitigation, and impact of misinformation

Problem

LLMs can generate highly convincing misinformation that exploits audience biases, and exposure was found to reduce trust and influence decision-making.

AI & SOCIETY
Gain

LLMs demonstrated capacity to detect false claims and increase users' resistance to misinformation, with personalized corrections showing effectiveness.

AI & SOCIETY
Engineering smart polymeric lipid nanoparticles for breast cancer: The convergence of AI and medical personalization to enhance efficacy and safety
HealthNegative state · G 65 / P 70

AI-designed smart polymeric-lipid nanoparticles for breast cancer treatment efficacy and safety

Source article: Engineering smart polymeric lipid nanoparticles for breast cancer: The convergence of AI and medical personalization to enhance efficacy and safety

Problem

No AI-designed polymeric-lipid nanoparticle formulation for breast cancer has entered registered clinical trials, leaving scalable production and regulatory clearance unresolved.

International Journal of Pharmaceutics
Gain

Preclinical studies indicate AI-optimized, ligand-functionalized polymeric-lipid nanoparticles can improve tumor-specific accumulation and reduce systemic side effects for breast cancer subtypes.

International Journal of Pharmaceutics
Unfolding new horizons: Machine learning applications for pediatric intussusception - A systematic review and meta-analysis
HealthPositive state · G 80 / P 75

machine learning models for pediatric intussusception diagnosis and prognosis

Source article: Unfolding new horizons: Machine learning applications for pediatric intussusception - A systematic review and meta-analysis

Problem

Review detected possible publication bias, found no studies reporting clinical endpoint data, and found abdominal radiograph triage models had substantially lower external accuracy, limiting readiness for clinical implementation.

Journal of Pediatric Surgery
Gain

Systematic review of 11 studies (36,863 patients) found ultrasound-based ML models achieved high pooled sensitivity and specificity with external validation, and AI assistance improved junior readers' specificity and cut examination time by 61% without loss of accuracy.

Journal of Pediatric Surgery
Guideline-augmented prompting improves comparative preference and response consistency of large language model outputs for orthopaedic anaesthesia questions: A controlled prompting study
HealthContested · G 71 / P 69

LLM-generated answers to orthopaedic anaesthesia questions evaluated for preference and consistency

Source article: Guideline-augmented prompting improves comparative preference and response consistency of large language model outputs for orthopaedic anaesthesia questions: A controlled prompting study

Problem

LLM responses for orthopaedic anaesthesia showed variable consistency, with later models exhibiting greater partial variability despite avoiding fully contradictory outputs.

Journal of Experimental Orthopaedics
Gain

Guideline-augmented prompting increased blinded preference win rates and improved response consistency for LLMs answering orthopaedic anaesthesia questions compared to human experts.

Journal of Experimental Orthopaedics
Development and Nationwide Multicentre Evaluation of Guideline-Grounded Large Language Model Chatbots to Support Patient Self-Management and Education in Rheumatology
HealthContested · G 67 / P 70

quality and safety of guideline-grounded rheumatology chatbot responses for patient self-management

Source article: Development and Nationwide Multicentre Evaluation of Guideline-Grounded Large Language Model Chatbots to Support Patient Self-Management and Education in Rheumatology

Problem

Chatbots failed to answer 4.2% of interactions, received negative ratings for insufficient detail, and showed only 45% full guideline adherence with weak agreement between LLM and physician assessments.

Journal of Medical Systems
Gain

Patients using guideline-grounded rheumatology chatbots reported high usability and satisfaction, with most answers rated safe and correct in real-world deployment.

Journal of Medical Systems
[Atrial fibrillation screening-reading the P-wave]
HealthContested · G 71 / P 69

AI-based P-wave risk scoring from sinus-rhythm ECG to guide atrial fibrillation screening and prolonged monitoring decisions

Source article: [Atrial fibrillation screening-reading the P-wave]

Problem

An AI-based P-wave risk score derived from sinus-rhythm ECG does not establish an AF diagnosis and alone does not constitute an indication for treatment, risking overinterpretation in screening.

Die Innere Medizin
Gain

AI models analyzing sinus-rhythm ECGs can identify atrial risk phenotypes and help select patients for prolonged rhythm monitoring to improve AF detection yield.

Die Innere Medizin
Generative AI versus physicians in diagnostic radiology: a systematic review and meta-analysis
HealthContested · G 71 / P 73

diagnostic accuracy of generative AI for radiology tasks

Source article: Generative AI versus physicians in diagnostic radiology: a systematic review and meta-analysis

Problem

Generative AI showed significantly lower diagnostic accuracy than expert physicians on radiology tasks, with a 13.0 percentage point gap.

Japanese Journal of Radiology
Gain

Generative AI achieved higher diagnostic accuracy when given text-only input compared to image-only input in radiology tasks.

Japanese Journal of Radiology
Explainable machine learning for identifying mild cognitive impairment in older adults with chronic diseases: Model development and temporal validation
HealthContested · G 69 / P 69

identifying current MCI status among older adults with chronic conditions using prespecified XGBoost model

Source article: Explainable machine learning for identifying mild cognitive impairment in older adults with chronic diseases: Model development and temporal validation

Problem

Model had limited sensitivity and calibration, with sensitivity dropping to 0.390 in temporal validation and Brier score worsening to 0.137, limiting clinical utility.

Acta Psychologica
Gain

XGBoost model achieved moderate and relatively stable discrimination for identifying current MCI status among older adults with chronic conditions across 2018 internal and 2020 temporal validation, with highly stable TreeSHAP explanations.

Acta Psychologica
Rule-Based Versus Generative Extraction of Psychological Symptoms From Forensic Medical Certificates: A Feasibility Study in the French ORFeAD Network
HealthNegative state · G 66 / P 71

automated extraction of psychological symptoms from French ORFeAD forensic medical certificates for interpersonal violence victims

Source article: Rule-Based Versus Generative Extraction of Psychological Symptoms From Forensic Medical Certificates: A Feasibility Study in the French ORFeAD Network

Problem

Sensitivity was heterogeneous and low for low-prevalence symptoms, 11 of 35 variables failed the reliability threshold under either pipeline, and neither pipeline supports individual-level decisions, with the generative model incurring far higher compute cost.

Behavioral Sciences & the Law
Gain

Automated extraction of psychological and subjective variables from unstructured forensic certificates was feasible with high specificity, with 24 of 35 variables meeting an 85% reliability threshold under at least one pipeline.

Behavioral Sciences & the Law