Deep-learning NLP models classified antidepressant treatment response as improved versus no evidence of improvement from routine EHR clinical notes with AUROC up to 0.88.
Researchers evaluated eight deep-learning natural language processing models to phenotype antidepressant treatment response from routine clinical notes in the Mass General Brigham system. Using 111,572 patients from 1990-2018 and 4,299 manually reviewed note sets across 2 days to 26 weeks after initiation, models distinguished 'improved' versus 'no evidence of improvement' with strong discrimination.
- Impact 30%
- 49
- Evidence 25%
- 95
- Scale 20%
- 35
- Confidence 15%
- 87
- Recency 10%
- 84
Updated Jul 17, 2026 · TRV-2026-0250
ace