Detecting Fake News Over Online Social Media via Domain Reputations and Content Understanding

Kuai Xu; Feng Wang; Haiyan Wang; Bo Yang

doi:10.26599/TST.2018.9010139

AI Chat Paper

Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.

Chat more with AI

| Sign up

Browse by Subject

Search for peer-reviewed journals with full access.

Journals A - Z

About Us

Discover the SciOpen Platform and Achieve Your Research Goals with Ease.

About Us

Publish with Us

Support

Journals A - Z

About Us

Publish with Us

Support

PDF (1.8 MB)

Cite

EndNote(RIS) BibTeX

Collect

Submit Manuscript

AI Chat Paper

Show Outline

Outline

Show full outline

Hide outline

Outline

Show full outline

Hide outline

Open Access

Detecting Fake News Over Online Social Media via Domain Reputations and Content Understanding

Kuai Xu(

), Feng Wang, Haiyan Wang, Bo Yang

School of Mathematical and Natural Sciences, Arizona State University, Glendale, AZ 85306, USA.

Jiangxi University of Finance and Economics, Nanchang 330013, China.

Show Author Information

Abstract

Fake news has recently leveraged the power and scale of online social media to effectively spread misinformation which not only erodes the trust of people on traditional presses and journalisms, but also manipulates the opinions and sentiments of the public. Detecting fake news is a daunting challenge due to subtle difference between real and fake news. As a first step of fighting with fake news, this paper characterizes hundreds of popular fake and real news measured by shares, reactions, and comments on Facebook from two perspectives: domain reputations and content understanding. Our domain reputation analysis reveals that the Web sites of the fake and real news publishers exhibit diverse registration behaviors, registration timing, domain rankings, and domain popularity. In addition, fake news tends to disappear from the Web after a certain amount of time. The content characterizations on the fake and real news corpus suggest that simply applying term frequency-inverse document frequency (tf-idf) and Latent Dirichlet Allocation (LDA) topic modeling is inefficient in detecting fake news, while exploring document similarity with the term and word vectors is a very promising direction for predicting fake and real news. To the best of our knowledge, this is the first effort to systematically study domain reputations and content characteristics of fake and real news, which will provide key insights for effectively detecting fake news on social media.

Keywords

social media fake news detection content modeling domain reputations

References

[1]

K. Shu, A. Sliva, S. H. Wang, J. L. Tang, and H. Liu, Fake news detection on social media: A data mining perspective, ACM SIGKDD Explorati., vol. 19, no. 1, pp. 22-36, 2017.

Crossref Google Scholar

[2]

E. Tacchini, G. Ballarin, M. L. D. Vedova, S. Moret, and L. de Alfaro, Some like it hoax: Automated fake news detection in social networks, Tech. Rep. UCSC-SOE-17-05, School of Engineering, University of California, Santa Cruz, CA, USA, 2017.

Google Scholar

[3]

M. M. Waldrop, News feature: The genuine problem of fake news, Proc. Natl. Acad. Sci. USA, vol. 114, no. 48, pp. 12631-12634, 2017.

Crossref Google Scholar

[4]

Z. B. He, Z. P. Cai, and X. M. Wang, Modeling propagation dynamics and developing optimized countermeasures for rumor spreading in online social networks, in Proc. IEEE 35th Int. Conf. on Distributed Computing Systems, Columbus, OH, USA, 2015.

Crossref

[5]

Z. B. He, Z. P. Cai, J. G. Yu, X. M. Wang, Y. C. Sun, and Y. S. Li, Cost-efficient strategies for restraining rumor spreading in mobile social networks, IEEE Trans. Veh. Technol., vol. 66, no. 3, pp. 2789-2800, 2017.

Crossref Google Scholar

[6]

C. C. Shao, G. L. Ciampaglia, O. Varol, A. Flammini, and F. Menczer, The spread of fake news by social bots, arXiv preprint arXiv: 1707.07592v1, 2017.

Google Scholar

[7]

J. Thorne, M. J. Chen, G. Myrianthous, J. S. Pu, X. X. Wang, and A. Vlachos, Fake news detection using stacked ensemble of classifiers, in Proc. EMNLP Workshop on Natural Language Processing Meets Journalism, Copenhagen, Denmark, 2017.

Crossref

[8]

N. J. Conroy, V. L. Rubin, and Y. M. Chen, Automatic deception detection: Methods for finding fake news, Proc. Assoc. Inf. Sci. Technol., vol. 52, no. 1, pp. 1-4, 2015.

Crossref Google Scholar

[9]

B. Markines, C. Cattuto, and F. Menczer, Social spam detection, in Proc. 5th Int. Workshop on Adversarial Information Retrieval on the Web, Madrid, Spain, 2009.

Crossref

[10]

Y. M. Chen, N. J. Conroy, and V. L. Rubin, Misleading online content: Recognizing clickbait as “false news”, in Proc. 2015 ACM on Workshop on Multimodal Deception Detection, Seattle, WA, USA, 2015.

Crossref

[11]

N. Ruchansky, S. Seo, and Y. Liu, CSI: A hybrid deep model for fake news detection, in Proc. 2017 ACM on Conf. on Information and Knowledge Management, Singapore, 2017.

[12]

BuzzFeed News, This analysis shows how viral fake election news stories outperformed real news On Facebook, https://www.buzzfeednews.com/article/craigsilverman/viral-fake-election-news-outperformed-real-news-on-facebook, 2016.

[13]

Alexa, https://www.alexa.com/topsites.

[14]

G. Salton and M. J. McGill, Introduction to Modern Information Retrieval. New York, NY, USA: McGraw-Hill, Inc, 1986.

[15]

D. M. Blei, A. Y. Ng, and M. I. Jordan, Latent Dirichlet allocation, J. Mach. Learn. Res., vol. 3, no. 3, pp. 993-1022, 2003.

Crossref Google Scholar

[16]

K. Stevens, P. Kegelmeyer, D. Andrzejewski, and D. Buttler, Exploring topic coherence over many models and many topics, in Proc. Joint Conf. on Empirical Methods in Natural Language Processing and Computational Natural Language Learning, Jeju Island, Korea, 2012.

Crossref

[17]

F. Wang, H. Y. Wang, K. Xu, J. H. Wu, and X. H. Jia, Characterizing information diffusion in online social networks with linear diffusive model, in Proc. IEEE Int. Conf. on Distributed Computing Systems, Philadelphia, PA, USA, 2013.

Crossref

[18]

V. L. Rubin, N. J. Conroy, Y. M. Chen, and S. Cornwell, Fake news or truth? Using satirical cues to detect potentially misleading news, in Proc. Ann. Conf. of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, San Diego, CA, USA, 2016.

Crossref

[19]

M. Hardalov, I. Koychev, and P. Nakov, In search of credible news, in Proc. 17th Int. Conf. on Artificial Intelligence: Methodology, Systems, and Applications, Varna, Bulgaria, 2016.

[20]

M. Farajtabar, J. C. Yang, X. J. Ye, H. Xu, R. Trivedi, E. Khalil, S. Li, L. Song, and H. Y. Zha, Fake news mitigation via point process based intervention, in Proc. 34th Int. Conf. on Machine Learning, Sydney, Australia, 2017.

[21]

C. Chen, K. Wu, V. Srinivasan, and X. D. Zhang, Battling the internet water army: Detection of hidden paid posters, in Proc. IEEE/ACM Int. Conf. on Advances in Social Networks Analysis and Mining, Niagara, Canada, 2013.

Crossref

[22]

F. Wang, K. Orton, P. Wagenseller, and K. Xu, Towards understanding community interests with topic modeling, IEEE Acc., vol. 6, pp. 24660-24668, 2018.

Crossref Google Scholar

[23]

K. Lei, Y. Liu, S. R. Zhong, Y. B. Liu, K. Xu, Y. Shen, and M. Yang, Understanding user behavior in Sina Weibo online social network: A community approach, IEEE Acc., vol. 6, pp. 13302-13316, 2018.

Crossref Google Scholar

[24]

K. Xu, F. Wang, H. Y. Wang, and B. Yang, A first step towards combating fake news over online social media, in Proc. 13th Int. Conf. on Wireless Algorithms, Systems, and Applications, Tianjin, China, 2018.

Crossref

[25]

V. L. Rubin, Y. M. Chen, and N. J. Conroy, Deception detection for news: Three types of fakes, in Proc. 78th ASIS&T Ann. Meeting: Information Science with Impact: Research in and for the Community, Saint Louis, MO, USA, 2015.

Crossref

[26]

W. Y. Wang, “Liar, liar pants on fire”: A new benchmark dataset for fake news detection, in Proc. 55th Ann. Meeting of the Association for Computational Linguistics, Vancouver, Canada, 2017.

Crossref

[27]

T. Mikolov, K. Chen, G. S. Corrado, and J. Dean, Efficient estimation of word representations in vector space, in Proc. Int. Conf. on Learning Representations, Scottsdale, AZ, USA, 2013.

Tsinghua Science and Technology

Volume 25 Issue 1,
February 2020

Pages 20-27

DOI: 10.26599/TST.2018.9010139

Cite this article:

Xu K, Wang F, Wang H, et al. Detecting Fake News Over Online Social Media via Domain Reputations and Content Understanding. Tsinghua Science and Technology, 2020, 25(1): 20-27. https://doi.org/10.26599/TST.2018.9010139

929

Views

124

Downloads

Crossref

N/A

Web of Science

Scopus

CSCD

Google Scholar
Citation

Altmetrics

Received: 07 October 2018

Revised: 10 November 2018

Accepted: 15 November 2018

Published: 22 July 2019

The articles published in this open access journal are distributed under the terms of the Creative Commons Attribution 4.0 International License (http://creativecommons.org/licenses/by/4.0/).