Cross-Language Fake News Detection

Samuel Kai Wah Chu, Runbin Xie, Yanshu Wang

Research output: Contribution to journalArticlepeer-review

24 Citations (Scopus)

Abstract

With increasing globalization, news from different countries, and even in different languages, has become readily available and has become a way for many people to learn about other cultures. As people around the world become more reliant on social media, the impact of fake news on public society also increases. However, most of the fake news detection research focuses only on English. In this work, we compared the difference between textual features of different languages (Chinese and English) and their effect on detecting fake news. We also explored the cross-language transmissibility of fake news detection models. We found that Chinese textual features in fake news are more complex compared with English textual features. Our results also illustrated that the bidirectional encoder representations from transformers (BERT) model outperformed other algorithms for within-language data sets. As for detection in cross-language data sets, our findings demonstrated that fake news monitoring across languages is potentially feasible, while models trained with data from a more inclusive language would perform better in cross-language detection.

Original languageEnglish
Pages (from-to)100-109
Number of pages10
JournalData and Information Management
Volume5
Issue number1
DOIs
Publication statusPublished - 1 Jan 2021
Externally publishedYes

Keywords

  • cross-language study
  • fake news detection
  • information detection

Fingerprint

Dive into the research topics of 'Cross-Language Fake News Detection'. Together they form a unique fingerprint.

Cite this