Assessing prognosis in depression: comparing perspectives of AI models, mental health professionals and the general public-Reference-Cited by-同舟云学术

Assessing prognosis in depression: comparing perspectives of AI models, mental health professionals and the general public

Published:2024-01 Issue:Suppl 1 Volume:12 Page:e002583
ISSN:2305-6983
Container-title:Family Medicine and Community Health
language:en
Short-container-title:Fam Med Com Health

Author:

Elyoseph Zohar,Levkovich Inbar^ORCID,Shinan-Altman Shiri

Abstract

BackgroundArtificial intelligence (AI) has rapidly permeated various sectors, including healthcare, highlighting its potential to facilitate mental health assessments. This study explores the underexplored domain of AI’s role in evaluating prognosis and long-term outcomes in depressive disorders, offering insights into how AI large language models (LLMs) compare with human perspectives.MethodsUsing case vignettes, we conducted a comparative analysis involving different LLMs (ChatGPT-3.5, ChatGPT-4, Claude and Bard), mental health professionals (general practitioners, psychiatrists, clinical psychologists and mental health nurses), and the general public that reported previously. We evaluate the LLMs ability to generate prognosis, anticipated outcomes with and without professional intervention, and envisioned long-term positive and negative consequences for individuals with depression.ResultsIn most of the examined cases, the four LLMs consistently identified depression as the primary diagnosis and recommended a combined treatment of psychotherapy and antidepressant medication. ChatGPT-3.5 exhibited a significantly pessimistic prognosis distinct from other LLMs, professionals and the public. ChatGPT-4, Claude and Bard aligned closely with mental health professionals and the general public perspectives, all of whom anticipated no improvement or worsening without professional help. Regarding long-term outcomes, ChatGPT 3.5, Claude and Bard consistently projected significantly fewer negative long-term consequences of treatment than ChatGPT-4.ConclusionsThis study underscores the potential of AI to complement the expertise of mental health professionals and promote a collaborative paradigm in mental healthcare. The observation that three of the four LLMs closely mirrored the anticipations of mental health experts in scenarios involving treatment underscores the technology’s prospective value in offering professional clinical forecasts. The pessimistic outlook presented by ChatGPT 3.5 is concerning, as it could potentially diminish patients’ drive to initiate or continue depression therapy. In summary, although LLMs show potential in enhancing healthcare services, their utilisation requires thorough verification and a seamless integration with human judgement and skills.

Publisher

BMJ

Subject

Family Practice,Public Health, Environmental and Occupational Health

Reference56 articles.

1. A systematic literature review of artificial intelligence in the healthcare sector: benefits, challenges, methodologies, and functionalities;Ali;Journal of Innovation & Knowledge,2023

2. Mariani MM , Machado I , Nambisan S . Types of innovation and artificial intelligence: a systematic quantitative literature review and research agenda. Journal of Business Research 2023;155:113364. doi:10.1016/j.jbusres.2022.113364

3. Chatgpt outperforms humans in emotional awareness evaluations;Elyoseph;Front Psychol,2023

4. Hadar-Shoval D , Elyoseph Z , Lvovsky M . The plasticity of ChatGPT's mentalizing abilities: personalization for personality structures. Front Psychiatry 2023;14:1234397. doi:10.3389/fpsyt.2023.1234397

5. Elyoseph Z , Levkovich I . Beyond human expertise: the promise and limitations of ChatGPT in suicide risk assessment. Front Psychiatry 2023;14:1213141. doi:10.3389/fpsyt.2023.1213141