Risks from Language Models for Automated Mental Healthcare: Ethics and Structure for Implementation
Grabb, D.; Lamparth, M.; Vasan, N.
Show abstract
Amidst the growing interest in developing task-autonomous AI for automated mental health care, this paper addresses the ethical and practical challenges associated with the issue and proposes a structured framework that delineates levels of autonomy, outlines ethical requirements, and defines beneficial default behaviors for AI agents in the context of mental health support. We also evaluate ten state-of-the-art language models using 16 mental health-related questions designed to reflect various mental health conditions, such as psychosis, mania, depression, suicidal thoughts, and homicidal tendencies. The question design and response evaluations were conducted by mental health clinicians (M.D.s). We find that existing language models are insufficient to match the standard provided by human professionals who can navigate nuances and appreciate context. This is due to a range of issues, including overly cautious or sycophantic responses and the absence of necessary safeguards. Alarmingly, we find that most of the tested models could cause harm if accessed in mental health emergencies, failing to protect users and potentially exacerbating existing symptoms. We explore solutions to enhance the safety of current models. Before the release of increasingly task-autonomous AI systems in mental health, it is crucial to ensure that these models can reliably detect and manage symptoms of common psychiatric disorders to prevent harm to users. This involves aligning with the ethical framework and default behaviors outlined in our study. We contend that model developers are responsible for refining their systems per these guidelines to safeguard against the risks posed by current AI technologies to user mental health and safety. Trigger warningContains and discusses examples of sensitive mental health topics, including suicide and self-harm.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Understanding Psychiatric Illness Through Natural Language Processing (UNDERPIN): Rationale, Design, and Methodology 90%
- Computational Psychiatry Research Map (CPSYMAP): a New Database for Visualizing Research Papers 90%
- Deep Multimodal Representations and Classification of First-Episode Psychosis via Live Face Processing 90%
Similar papers in this journal
Similar papers in this journal
- Listening to mental health crisis needs at scale: using Natural Language Processing to understand and evaluate a mental health crisis text messaging service 95%
- Large Language Models in Real-World Clinical Workflows: A Systematic Review of Applications and Implementation 91%
- The development of a World Health Organization transdiagnostic chatbot intervention for distressed adolescents and young adults 89%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.