Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Zaihan Yang

Detection of Suicidal Risk on Social Media: A Hybrid Model

May 26, 2025

Zaihan Yang, Ryan Leonard, Hien Tran, Rory Driscoll, Chadbourne Davis

Abstract:Suicidal thoughts and behaviors are increasingly recognized as a critical societal concern, highlighting the urgent need for effective tools to enable early detection of suicidal risk. In this work, we develop robust machine learning models that leverage Reddit posts to automatically classify them into four distinct levels of suicide risk severity. We frame this as a multi-class classification task and propose a RoBERTa-TF-IDF-PCA Hybrid model, integrating the deep contextual embeddings from Robustly Optimized BERT Approach (RoBERTa), a state-of-the-art deep learning transformer model, with the statistical term-weighting of TF-IDF, further compressed with PCA, to boost the accuracy and reliability of suicide risk assessment. To address data imbalance and overfitting, we explore various data resampling techniques and data augmentation strategies to enhance model generalization. Additionally, we compare our model's performance against that of using RoBERTa only, the BERT model and other traditional machine learning classifiers. Experimental results demonstrate that the hybrid model can achieve improved performance, giving a best weighted $F_{1}$ score of 0.7512.

Via

Access Paper or Ask Questions

Detecting People Interested in Non-Suicidal Self-Injury on Social Media

Jul 10, 2022

Zaihan Yang, Dmitry Zinoviev

Figure 1 for Detecting People Interested in Non-Suicidal Self-Injury on Social Media

Figure 2 for Detecting People Interested in Non-Suicidal Self-Injury on Social Media

Abstract:We propose a supervised learning approach to detect people interested in Non-Suicidal Self-Injury (NSSI). We treat the task as a binary classification problem, and build classifiers based upon features extracted from people self-declared interests. Experimental evaluation on a real-world dataset, the LiveJournal social blogging networking platform, demonstrates the effectiveness of our proposed model.

* The 9th International Conference on Complex Networks and Their Applications - Book of Abstracts, pp.414 (978-2-9557050-4-9) 2021
* 3 pages

Via

Access Paper or Ask Questions