Comparative analysis of machine learning methods to detect fake news in an Urdu language .

Advanced Search

Adnan Rafique, Furqan Rustam, Manideep Narra, Arif Mehmood, Ernesto Lee, Imran Ashraf

Author Information

Adnan Rafique: Department of Computer Science, COMSATS University Islamabad (CUI), Lahore, Pakistan. ORCID
Furqan Rustam: Department of Software Engineering, University of Management and Technology, Lahore, Pakistan. ORCID
Manideep Narra: Indiana Institute of Technology, Fort Wayne, United States. ORCID
Arif Mehmood: Department of CS and IT, Islamia University, Bahawalpur, Bahawalpur, Pakistan.
Ernesto Lee: School of Engineering and Technology, Miami Dade College, Miami, FL, USA. ORCID
Imran Ashraf: Information and Communication Engineering, Yeungnam University, Gyeongsan si, Daegu, South Korea. ORCID

PMID: 35875651 DOI: 10.7717/peerj-cs.1004

Wide availability and large use of social media enable easy and rapid dissemination of news. The extensive spread of engineered news with intentionally false information has been observed over the past few years. Consequently, fake news detection has emerged as an important research area. Fake news detection in the Urdu language spoken by more than 230 million people has not been investigated very well. This study analyzes the use and efficacy of various machine learning classifiers along with a deep learning model to detect fake news in the Urdu language. Logistic regression, support vector machine, random forest (RF), naive Bayes, gradient boosting, and passive aggression have been utilized to this end. The influence of term frequency-inverse document frequency and BoW features has also been investigated. For experiments, a manually collected dataset that contains 900 news articles was used. Results suggest that RF performs better and achieves the highest accuracy of 0.92 for Urdu fake news with BoW features. In comparison with machine learning models, neural networks models long short term memory, and multi-layer perceptron are used. Machine learning models tend to show better performance than deep learning models.

Deep learning Fake news detection Machine learning Urdu corpus

Front Neurorobot. 2013 Dec 04;7:21 [PMID: 24409142]
J Trauma. 1987 Apr;27(4):370-8 [PMID: 3106646]

News

OpenLB
Open Library of Bioscience