Mental Health Disorder Indications Detection Based on Text Using NLP with EDA Augmentation Techniques

Authors

  • Erna Daniati Universitas Nusantara PGRI Kediri
  • Sherly Dian Tiara Universitas Nusantara PGRI Kediri
  • Arie Nugroho Universitas Nusantara PGRI Kediri

DOI:

https://doi.org/10.59095/ijcsr.v5i2.273

Keywords:

Early Detection, Easy Data Augmentation, Logistic Regression, Natural Language Processing, Mental Health

Abstract

This study aims to build a classification model for the early screening of mental health disorders from social media text data using the CRISP-DM framework. The primary issue of data imbalance between categories was addressed using the Easy Data Augmentation (EDA) technique. Logistic Regression algorithm and TF-IDF feature extraction were used to classify six categories of mental conditions. Test results showed that the model with EDA experienced a slight decrease in global accuracy to 0.74 (compared to 0.76 without EDA) but successfully increased the Recall for the minority class, Mentalillness, significantly from 0.28 to 0.56. This improvement proves that EDA effectively enriches linguistic variation in limited data. The model has been validated by a psychologist and implemented into a web-based application as an indicative early detection tool, not a clinical medical diagnosis.

 

Downloads

Download data is not yet available.

Downloads

Published

2026-07-31

Issue

Section

Articles

How to Cite

Mental Health Disorder Indications Detection Based on Text Using NLP with EDA Augmentation Techniques. (2026). The Indonesian Journal of Computer Science Research, 5(2), 108-115. https://doi.org/10.59095/ijcsr.v5i2.273

Similar Articles

31-40 of 48

You may also start an advanced similarity search for this article.