Vol. 29 No. 3 (2026) Cover Image
Vol. 29 No. 3 (2026)

Published: September 20, 2026

Pages: 521-529

Articles

Enhanced Machine Learning Framework for Detecting Fake News Used as Soft Power in Textual Content

Abstract

The utilisation of fake news as a tool of soft power is regarded as a strategic component of contemporary warfare, posing an escalating threat to societal and political stability as well as public trust. Deep Learning models achieve superior performance; yet, they require substantial computational resources and possess interpretative restrictions, which hinder their practical application in real-time scenarios and impose constraints on resource configurations. This study evaluates a traditional machine-learning framework based on linear and ensemble classifiers, including Linear Support Vector Classification (LinearSVC), Logistic Regression, Multinomial Naïve Bayes, and Random Forest. The framework employs a fixed, explicitly specified TF–IDF vectorization configuration and a consistent preprocessing pipeline. The system was assessed using the WELFake dataset, which contains over 70,000 labelled news stories, employing accuracy, precision, recall, F1-score, and ROC-AUC as evaluation metrics. The findings demonstrate the superiority of LinearSVC, attaining the maximum accuracy of 95.65%, providing balanced performance metrics, and surpassing all other models. This study contributes by providing an interpretable, scalable, and domain-agnostic solution through the development of a practical fake news detection system designed to combat digital disinformation in real-world contexts.

References

  1. B. Rød, C. Pursiainen, and N. Eklund, "Combatting Disinformation – How Do We Create Resilient Societies? Literature Review and Analytical Framework," European Journal for Security Research, 2025, https://doi.org10.1007/s41125-025-00105-4
  2. P. N. Vasist and S. Krishnan, "Country branding in post-truth Era: A configural narrative," Journal of Destination Marketing & Management, vol. 32, p. 100854, 2024, https://doi.org/10.1016/j.jdmm.2024.100854
  3. J. Arayankalam and S. Krishnan, "Relating foreign disinformation through social media, domestic online media fractionalization, government's control over cyberspace, and social media-induced offline violence: Insights from the agenda-building theoretical perspective," Technological Forecasting and Social Change, vol. 166, p. 120661, 2021, https://doi.org/10.1016/j.techfore.2021.120661
  4. R. Marigliano, L. H. X. Ng, and K. M. Carley, "Analyzing digital propaganda and conflict rhetoric: a study on Russia’s bot-driven campaigns and counter-narratives during the Ukraine crisis," Social Network Analysis and Mining, vol. 14, no. 1, 2024, https://doi.org10.1007/s13278-024-01322-w
  5. M. Soprano et al., "Cognitive Biases in Fact-Checking and Their Countermeasures: A Review," Information Processing & Management, vol. 61, no. 3, p. 103672, 2024, https://doi.org10.1016/j.ipm.2024.103672
  6. R. Ș. Emil and B. Remus, "A Review of Automatic Fake News Detection: From Traditional Methods to Large Language Models," Future Internet, vol. 17, no. 10, p. 435, 2025, https://doi.org10.3390/fi17100435
  7. E. S. Albtoush, K. H. Gan, and S. A. A. Alrababa, "Fake news detection: state-of-the-art review and advances with attention to Arabic language aspects," PeerJ Computer Science, vol. 11, p. e2693, 2025, https://doi.org10.7717/peerj-cs.2693
  8. J. Alghamdi, Y. Lin, and S. Luo, "A Comparative Study of Machine Learning and Deep Learning Techniques for Fake News Detection," Information, vol. 13, no. 12, p. 576, 2022. [Online]. Available: https://www.mdpi.com/2078-2489/13/12/576
  9. P. Linardatos, V. Papastefanopoulos, and S. Kotsiantis, "Explainable AI: A Review of Machine Learning Interpretability Methods," Entropy, vol. 23, no. 1, p. 18, 2021. [Online]. Available: https://www.mdpi.com/1099-4300/23/1/18
  10. E. Strubell, A. Ganesh, and A. McCallum, "Energy and Policy Considerations for Deep Learning in NLP," Florence, Italy, July 2019: Association for Computational Linguistics, in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, pp. 3645-3650, https://doi.org10.18653/v1/P19-1355
  11. P. K. Verma, P. Agrawal, and R. Prodan. “WELFake dataset for fake news detection in text data”, Zenodo, 2021, https://doi.org10.5281/zenodo.4561253
  12. P. K. Verma, P. Agrawal, I. Amorim, and R. Prodan, "WELFake: Word Embedding Over Linguistic Features for Fake News Detection," IEEE Transactions on Computational Social Systems, vol. 8, no. 4, pp. 881-893, 2021, https://doi.org10.1109/TCSS.2021.3068519
  13. D. Dementieva, M. Kuimov, and A. Panchenko, "Multiverse: Multilingual Evidence for Fake News Detection," Journal of Imaging, vol. 9, no. 4, p. 77, 2023. https://www.mdpi.com/2313-433X/9/4/77 .
  14. R. Mohawesh, S. Maqsood, and Q. Althebyan, "Multilingual deep learning framework for fake news detection using capsule neural network," Journal of Intelligent Information Systems, vol. 60, no. 3, pp. 655-671, 2023, https://doi.org10.1007/s10844-023-00788-y
  15. D. K. Sharma and S. Garg, "IFND: a benchmark dataset for fake news detection," Complex & Intelligent Systems, vol. 9, no. 3, pp. 2843-2863, 2023, https://doi.org10.1007/s40747-021-00552-1
  16. T. Kumarage et al., "J-Guard: Journalism Guided Adversarially Robust Detection of AI-generated News," Nusa Dua, Bali, November 2023: Association for Computational Linguistics, in Proceedings of the 13th International Joint Conference on Natural Language Processing and the 3rd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics (Volume 1: Long Papers), pp. 484-497, https://doi.org10.18653/v1/2023.ijcnlp-main.32
  17. S. Muñoz and C. Á. Iglesias, "Exploiting Content Characteristics for Explainable Detection of Fake News," Big Data and Cognitive Computing, vol. 8, no. 10, p. 129, 2024. [Online]. Available: https://www.mdpi.com/2504-2289/8/10/129
  18. J. Alghamdi, Y. Lin, and S. Luo, "Enhancing hierarchical attention networks with CNN and stylistic features for fake news detection," Expert Systems with Applications, vol. 257, p. 125024, 2024, https://doi.orghttps://doi.org/10.1016/j.eswa.2024.125024
  19. M. E.Almandouh, M. F. Alrahmawy, M. Eisa, M. Elhoseny, and A. S. Tolba, "Ensemble based high performance deep learning models for fake news detection," Scientific Reports, vol. 14, no. 1, 2024, https://doi.org10.1038/s41598-024-76286-0
  20. S. Maham, A. Tariq, M. U. G. Khan, F. S. Alamri, A. Rehman, and T. Saba, "ANN: adversarial news net for robust fake news classification," Scientific Reports, vol. 14, no. 1, 2024, https://doi.org10.1038/s41598-024-56567-4
  21. R. Panchendrarajan and A. Zubiaga, "Claim detection for automated fact-checking: A survey on monolingual, multilingual and cross-lingual research," Natural Language Processing Journal, vol. 7, p. 100066, 2024, https://doi.orghttps://doi.org/10.1016/j.nlp.2024.100066
  22. G. Salton and C. Buckley, "Term-weighting approaches in automatic text retrieval," Information Processing & Management, vol. 24, no. 5, pp. 513-523, 1988/01/01/ 1988, https://doi.orghttps://doi.org/10.1016/0306-4573(88)90021-0
  23. S. Kapoor and A. Narayanan, "Leakage and the reproducibility crisis in machine-learning-based science," Patterns, vol. 4, no. 9, 2023, https://doi.org10.1016/j.patter.2023.100804
  24. P. Billion Polak, J. D. Prusa, and T. M. Khoshgoftaar, "Low-shot learning and class imbalance: a survey," Journal of Big Data, vol. 11, no. 1, p. 1, 2024, https://doi.org10.1186/s40537-023-00851-z
  25. K. Taha, P. D. Yoo, C. Yeun, D. Homouz, and A. Taha, "A comprehensive survey of text classification techniques and their research applications: Observational and experimental insights," Computer Science Review, vol. 54, p. 100664, 2024, https://doi.orghttps://doi.org/10.1016/j.cosrev.2024.100664
  26. C. Cortes and V. Vapnik, "Support-vector networks," Machine Learning, vol. 20, no. 3, pp. 273-297, 1995, https://doi.org10.1007/bf00994018
  27. G. James, D. Witten, T. Hastie, and R. Tibshirani, “An introduction to statistical learning: with applications in R”, Springer, 2013.
  28. C. D. Manning, P. Raghavan, and H. Schütze, Introduction to Information Retrieval. Cambridge: Cambridge University Press, 2008.
  29. L. Breiman, "Random Forests," Machine Learning, vol. 45, no. 1, pp. 5-32, 2001, https://doi.org10.1023/A:1010933404324
  30. G. Van Houdt, C. Mosquera, and G. Nápoles, "A review on the long short-term memory model," Artificial Intelligence Review, vol. 53, no. 8, pp. 5929-5955, 2020, https://doi.org10.1007/s10462-020-09838-1
  31. S. Hochreiter and J. Schmidhuber, "Long Short-Term Memory," Neural Computation, vol. 9, no. 8, pp. 1735-1780, 1997, https://doi.org10.1162/neco.1997.9.8.1735
  32. M. Sokolova and G. Lapalme, "A systematic analysis of performance measures for classification tasks," Information Processing & Management, vol. 45, no. 4, pp. 427-437,2009, https://doi.orghttps://doi.org/10.1016/j.ipm.2009.03.002