Back

Enhancing Mental Health Condition Detection on Social Media through Multi-Task Learning

Liu, J.; Su, M.

2024-02-27 health informatics
10.1101/2024.02.23.24303303 medRxiv
Show abstract

ObjectiveMental health conditions are traditionally modeled individually, which ignores the complex, interconnected nature of mental health disorders, which often share overlapping symptoms. This study aims to develop an integrated multi-task learning framework to enhance the detection of mental health conditions. MethodUtilizing datasets from Reddits SuicideWatch and Mental Health Collection (SWMH) and Psychiatric-disorder Symptoms (PsySym), the study develops a BERT-based multi-task learning framework. This framework leverages pre-trained embedding layers of BERT variants to capture linguistic nuances relevant to various mental health conditions from social media narratives. The approach is tested against the two datasets, comparing multitask modeling with a wide array of single-task baselines and large language models (LLM). ResultsThe multi-task learning framework demonstrated higher performance in efficiently predicting mental health conditions together compared to single-task models and general-purpose LLMs. Specifically, the framework achieved higher F1 scores across multiple conditions, with notable improvements in recall and precision metrics. This indicates more accurate modeling of mental health disorders when considered together, rather than in isolation. ConclusionThe study confirms the effectiveness of a multi-task learning approach in enhancing the detection of mental health conditions from social media data. It sets a new precedent in computational psychiatry and suggests future explorations into multi-task frameworks for deeper insights into mental health disorders.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.