Post

New research · AI in Medicine
NPJ digital medicine · 1w
AI / informaticsNPJ digital medicine · 2026

Evaluating large language models for assessment of psychosis risk.

Taiyu Zhu, Alexander Tashevski, Maxime Taquet … Dominic Oliver
Read paper
AI in MedicineAI / informatics

Large language models accurately identified clinical high risk for psychosis from interview transcripts

Evaluating large language models for assessment of psychosis risk.

Taiyu Zhu … Dominic Oliver
NPJ digital medicine · 2026
Background

Psychosis prevention relies on early detection of individuals at clinical high risk for psychosis (clinical high risk for psychosis).

Methods

We assessed 11 open-weight LLMs on 678 partial PSYCHS interview transcripts from 373 participants (77.7% clinical high risk for psychosis).

n = 373 participants
0.80
Results
0.80
the AI's calls on who was at high psychosis risk were mostly correct
n = 373 participants
More results

Models inferred clinical high risk for psychosis status and estimated severity and frequency across 15 symptom domains, benchmarked against researcher-rated scores.

LLM-generated symptom scores showed good correlations with researcher-rated scores (ICC sev = 0.74, ICC freq = 0.75).

More results

Generated summaries were largely faithful to source transcripts, with low rates of clinically relevant confabulation (3%).

While accuracy scaled with model size, smaller models achieved competitive performance with substantially lower computational cost.

“
Conclusion

These findings demonstrate that open-weight LLMs have the potential to assess psychosis risk from psychometric interview transcripts, supporting scalable, human-in-the-loop approaches to early detection.

Read paper
0 comments

No comments yet. Be the first.

Related papers

LatestFoundational
Cohort Study
3.79 points
Traditional Chinese medicine and acupuncture resulted in greater depression improvement than selective serotonin reuptake inhibitors
Guideline
“The CANMAT/ICOCS 2025 OCD International Guidelines synthesize the evidence on the efficacy, safety, and tolerability of the range of interventions available for the management of OCD.”
Study
OR 47.2
those with active mental illness symptoms were far more likely found unfit for trial
Cohort Study
Adjusted OR 1.65
higher blood-fat balance meant greater odds of depression after birth
Cohort Study
43.8%
nearly half of online mental-health care patients had both anxiety and depression together
Guideline
“It is hoped that the current Guidelines provide useful recommendations in important aspects of care for people living with schizophrenia and their family, whānau and carers in Australia and Aotearoa New Zealand, both at the individual and systemic levels.”
AI / Informatics
r = .70-.81
GPT-4's depression ratings closely matched standard questionnaires and expert clinicians' judgments
Cohort Study
AUC = 0.856
how well the model separated infants whose own liver survived 2 years from those who needed a transplant; higher means b