Evaluating the clinical safety of large language models in response to high-risk mental health disclosures

João M. Santos; Siddharth Shah; Amit Gupta; Aarav Mann; Alexandre Vaz; Benjamin E. Caldwell; Robert Scholz; Peter Awad; Rocky Allemandi; Doug Faust; Harshita Banka; Tony Rousmaniere

Ciência_Iscte Publicações Descrição Detalhada da Publicação

Artigo em revista científica

Evaluating the clinical safety of large language models in response to high-risk mental health disclosures

João M. Santos (Santos, J. M.); Siddharth Shah (Shah, S.); Amit Gupta (Gupta, A.); Aarav Mann (Mann, A.); Alexandre Vaz (Vaz, A.); Benjamin E. Caldwell (Caldwell, B. E.); Robert Scholz (Scholz, R.); Peter Awad (Awad, P.); Rocky Allemandi (Allemandi, R.); Doug Faust (Faust, D.); Harshita Banka (Banka, H.); Tony Rousmaniere (Rousmaniere, T.); et al.

Título Revista

Practice Innovations

Ano (publicação definitiva)

N/A

Língua

Inglês

País

Estados Unidos da América

Mais Informação

Visitar Link

Web of Science®

N.º de citações: 0

(Última verificação: 2026-05-29 02:59)

Ver o registo na Web of Science®

Scopus

Esta publicação não está indexada na Scopus

Google Scholar

N.º de citações: 0

(Última verificação: 2026-05-27 18:28)

Ver o registo no Google Scholar

Overton

Esta publicação não está indexada no Overton

Abstract/Resumo

As large language models increasingly mediate emotionally sensitive conversations, especially in mental health contexts, their ability to recognize and respond to high-risk situations becomes a matter of public safety. This study evaluates the responses of six popular large language models—Claude, Gemini, DeepSeek, ChatGPT, Grok 3, and LLAMA—to user prompts simulating crisis-level mental health disclosures. Drawing on a coding framework developed by licensed clinicians, five safety-oriented behaviors were assessed: explicit risk acknowledgment, empathy, encouragement to seek help, provision of specific resources, and invitation to continue the conversation. Claude outperformed all others in a global assessment, while Grok 3, ChatGPT, and LLAMA underperformed across multiple domains. Notably, most models exhibited empathy, but few consistently provided practical support or kept the conversation open. These findings suggest that while large language models show potential for emotionally attuned communication, none currently meet satisfactory clinical standards for crisis response. Ongoing development and targeted fine-tuning are essential to ensure ethical deployment of AI in mental health settings.

Agradecimentos/Acknowledgements

Palavras-chave

Large language models,Crisis intervention,Ethics,Mental health

Identificadores da Publicação

WoS (fonte: Ciência_Iscte)	RC:164925880_S24
DOI (fonte: autor)	10.1037/pri0000316
Handle (fonte: Ciência-IUL)	http://hdl.handle.net/10071/36775
DOI (fonte: outro)	10.1037/pri0000316
WoS (fonte: autor)	RC:164925880_S24
Outro ID (fonte: Externo)	cv-prod-id-5154095
ID Ciência_Iscte	ci-pub-117447

Outros Detalhes da Publicação

Ano Publicação Online	2026
Editora	American Psychological Association
Indexação	Web of Science©;
ISSN	2377-889X (print) 2377-8903 (online)
ISBN	--
Factor de Impacto	--
Volume	N/A	Número
Série
Número Artigo
Páginas	--
Avaliado Cientificamente	Sim
Repositório ISCTE-IUL	Link para o repositório
Data Publicação (online)	2026-02-16
Data Publicação (print)

Altmetric

Dimensions

PlumX Metrics