Detection of Suicidal Risk on Social Media: A Hybrid Model

Yang, Zaihan; Leonard, Ryan; Tran, Hien; Driscoll, Rory; Davis, Chadbourne

Computer Science > Computation and Language

arXiv:2505.23797 (cs)

[Submitted on 26 May 2025]

Title:Detection of Suicidal Risk on Social Media: A Hybrid Model

Authors:Zaihan Yang, Ryan Leonard, Hien Tran, Rory Driscoll, Chadbourne Davis

View PDF HTML (experimental)

Abstract:Suicidal thoughts and behaviors are increasingly recognized as a critical societal concern, highlighting the urgent need for effective tools to enable early detection of suicidal risk. In this work, we develop robust machine learning models that leverage Reddit posts to automatically classify them into four distinct levels of suicide risk severity. We frame this as a multi-class classification task and propose a RoBERTa-TF-IDF-PCA Hybrid model, integrating the deep contextual embeddings from Robustly Optimized BERT Approach (RoBERTa), a state-of-the-art deep learning transformer model, with the statistical term-weighting of TF-IDF, further compressed with PCA, to boost the accuracy and reliability of suicide risk assessment. To address data imbalance and overfitting, we explore various data resampling techniques and data augmentation strategies to enhance model generalization. Additionally, we compare our model's performance against that of using RoBERTa only, the BERT model and other traditional machine learning classifiers. Experimental results demonstrate that the hybrid model can achieve improved performance, giving a best weighted $F_{1}$ score of 0.7512.

Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Computers and Society (cs.CY); Machine Learning (cs.LG); Social and Information Networks (cs.SI)
Cite as:	arXiv:2505.23797 [cs.CL]
	(or arXiv:2505.23797v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2505.23797

Submission history

From: Zaihan Yang [view email]
[v1] Mon, 26 May 2025 14:56:47 UTC (380 KB)

Computer Science > Computation and Language

Title:Detection of Suicidal Risk on Social Media: A Hybrid Model

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Detection of Suicidal Risk on Social Media: A Hybrid Model

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators