Skip to content
Research Article Open access CC BY 4.0

Online Disinformation Detection through Text Analysis: A Comparative Study of Supervised Models with Hyperparameter Optimization

Viviane Kaseka Katadi, Gilbert Ngoyi Maloba, Jean-Marie Ibanga Mbayo, Pélagie Mpembe Mukonkole, Clarisse Ngoyi Tshite, Mardochée Kalala Mpolesha, Pierre Kafunda Katalay

Asian Journal of Research in Computer Science · pp. 55–72 · Published 15 Apr 2026

10.9734/ajrcos/2026/v19i4849

Abstract

The rapid spread of disinformation on social media poses a major challenge in the digital age, with significant impacts on public opinion and decision-making. In this context, this study proposes a machine learning-based approach for the automatic detection of online disinformation. A comparative analysis is conducted on several supervised learning models, including logistic regression, support vector machines (SVMs), random forests, and gradient boosting. The experiment is based on a real-world dataset of textual content from digital platforms, preprocessed using TF-IDF. Furthermore, hyperparameter optimization, primarily using Grid Search, is implemented to improve model performance. The results obtained reveal very high performance for all models, with accuracy values ​​exceeding 98% and areas under the ROC curve (AUC) close to 1. The Gradient Boosting model stands out as the best performer, offering an excellent balance between accuracy and generalization capabilities, while the Random Forest model, although exhibiting a perfect AUC, shows potential signs of overfitting. This study highlights the effectiveness of machine learning methods for disinformation detection and underscores the importance of hyperparameter optimization in improving model performance. It also opens up interesting avenues for integrating more advanced techniques, including deep learning and multimodal analysis, into disinformation countermeasures systems. The models were evaluated using data separation into training and test sets, allowing for a reliable estimation of their performance. The results show that hyperparameter optimization significantly improves the performance of classical models. However, certain limitations related to the diversity of data sources and methodological choices must be taken into account. Graphical Summary  

Disinformation machine learning hyperparameters; TF-IDF supervised learning fake-news detection

Cited by 0

No indexed citations yet.

Article metrics

Real usage data collected on this platform.

0

Page views

0

PDF downloads

0

Outbound clicks

0

Citations

Views by country

Approximate, from request IP at view time — not citizenship or institution. Countries with fewer than 5 views are grouped as "Other".

No views recorded yet.

Traffic sources

Referring site, by host.

No traffic recorded yet.

Views and downloads exclude known bots/crawlers. Citations combines this platform's own DOI-resolved index with each external source's own reported total — see Cited by above for individually listed citing works. Last refreshed 0 seconds ago.