Interpretable fake news detection with topic and deep variational models

Online Social Networks and Media - Tập 36 - Trang 100249 - 2023
Marjan Hosseini1, Alireza Javadian Sabet2, Suining He1, Derek Aguiar1
1Department of Computer Science and Engineering, University of Connecticut, Storrs, 06269, CT, USA
2Department of Informatics and Networked Systems, University of Pittsburgh, Pittsburgh, 15260, PA, USA

Tài liệu tham khảo

Lazer, 2018, The science of fake news, Science, 359, 1094, 10.1126/science.aao2998 Wardle, 2017 Posetti, 2018, A short guide to the history of’fake news’ and disinformation, Int. Cent. Journal., 7, 1 Cinelli, 2021, The echo chamber effect on social media, Proc. Natl. Acad. Sci., 118, 10.1073/pnas.2023301118 M. Chalkiadakis, A. Kornilakis, P. Papadopoulos, E. Markatos, N. Kourtellis, The Rise and Fall of Fake News sites: A Traffic Analysis, in: 13th ACM Web Science Conference 2021, 2021, pp. 168–177. Allcott, 2017, Social media and fake news in the 2016 election, J. Econ. Perspect., 31, 211, 10.1257/jep.31.2.211 Vosoughi, 2018, The spread of true and false news online, Science, 359, 1146, 10.1126/science.aap9559 Society of Professional Journalists, 2021 Carminati, 2012, Trust and share: Trusted information sharing in online social networks, 1281 Hindman, 2018 Lee, 2019, The global rise of “fake news” and the threat to democratic elections in the USA, Public Adm. Policy Guess, 2020, “Fake news” may have limited effects beyond increasing beliefs in false claims, Harv. Kennedy School Misinformation Rev., 1 Clayton, 2020, Real solutions for fake news? Measuring the effectiveness of general warnings and fact-check tags in reducing belief in false stories on social media, Political Behav., 42, 1073, 10.1007/s11109-019-09533-0 M. Babaei, A. Chakraborty, J. Kulshrestha, E.M. Redmiles, M. Cha, K.P. Gummadi, Analyzing biases in perception of truth in news stories and their implications for fact checking, in: Proceedings of the Conference on Fairness, Accountability, and Transparency, 2019, pp. 139–139. Politifact, 2020 Snopes, 2019 Pennycook, 2019, Fighting misinformation on social media using crowdsourced judgments of news source quality, Proc. Natl. Acad. Sci., 116, 2521, 10.1073/pnas.1806781116 Došilović, 2018, Explainable artificial intelligence: A survey, 0210 Gilpin, 2018, Explaining explanations: An overview of interpretability of machine learning, 80 Rudin, 2019, Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead, Nat. Mach. Intell., 1, 206, 10.1038/s42256-019-0048-x D. Khattar, J.S. Goud, M. Gupta, V. Varma, Mvae: Multimodal variational autoencoder for fake news detection, in: The World Wide Web Conference, 2019, pp. 2915–2921. Li, 2014, Spotting fake reviews via collective positive-unlabeled learning, 899 Qian, 2018, Neural user response generator: Fake news detection with collective user intelligence, 3834 Wang, 2017 Zhang, 2020, Fakedetector: Effective fake news detection with deep diffusive neural network, 1826 Stieglitz, 2018, Social media analytics – Challenges in topic discovery, data collection, and data preparation, Int. J. Inf. Manage., 39, 156, 10.1016/j.ijinfomgt.2017.12.002 One, 2021 Moens, 2014 Ahmed, 2018, Detecting opinion spams and fake news using text classification, Secur. Priv., 1 Banik, 2020 Boididou, 2018, Detection and visualization of misleading content on Twitter, Int. J. Multimedia Inf. Retr., 7, 71, 10.1007/s13735-017-0143-x Mikolov, 2010, Recurrent neural network based language model, 1045 Antoun, 2020, State of the art models for fake news detection tasks, 519 Blei, 2003, Latent dirichlet allocation, J. Mach. Learn. Res., 3, 993 Kingma, 2013 Hochreiter, 1997, Long short-term memory, Neural Comput., 9, 1735, 10.1162/neco.1997.9.8.1735 Graves, 2005, Framewise phoneme classification with bidirectional LSTM and other neural network architectures, Neural Netw., 18, 602, 10.1016/j.neunet.2005.06.042 Liu, 2019, Bidirectional LSTM with attention mechanism and convolutional layer for text classification, Neurocomputing, 337, 325, 10.1016/j.neucom.2019.01.078 Mikolov, 2013 Cunningham, 2015, Linear dimensionality reduction: Survey, insights, and generalizations, J. Mach. Learn. Res., 16, 2859 Rosipal, 2001, Kernel PCA for feature extraction and de-noising in nonlinear regression, Neural Comput. Appl., 10, 231, 10.1007/s521-001-8051-z Ringnér, 2008, What is principal component analysis?, Nature Biotechnol., 26, 303, 10.1038/nbt0308-303 Van der Maaten, 2008, Visualizing data using t-SNE, J. Mach. Learn. Res., 9 Narayan, 2021, Assessing single-cell transcriptomic variability through density-preserving data visualization, Nature Biotechnol., 39, 765, 10.1038/s41587-020-00801-7 Jin, 2014, News credibility evaluation on microblog with a hierarchical propagation model, 230 Baly, 2018 Zhou, 2020, A survey of fake news: Fundamental theories, detection methods, and opportunities, ACM Comput. Surv., 53, 10.1145/3395046 Afroz, 2012, Detecting hoaxes, frauds, and deception in writing style online, 461 H. Rashkin, E. Choi, J.Y. Jang, S. Volkova, Y. Choi, Truth of varying shades: Analyzing language in fake news and political fact-checking, in: Proc. of the 2017 Conference on Empirical Methods in Natural Language Processing, 2017, pp. 2931–2937. V.L. Rubin, N. Conroy, Y. Chen, S. Cornwell, Fake news or truth? using satirical cues to detect potentially misleading news, in: Proceedings of the Second Workshop on Computational Approaches To Deception Detection, 2016, pp. 7–17. Liu, 2010, Sentiment analysis and subjectivity, Handb. Nat. Lang. Process., 2, 627 Yadollahi, 2017, Current state of text sentiment analysis from opinion to emotion mining, ACM Comput. Surv., 50, 1, 10.1145/3057270 J. Ito, J. Song, H. Toda, Y. Koike, S. Oyama, Assessment of tweet credibility with LDA features, in: Proceedings of the 24th WWW, 2015, pp. 953–958. Ma, 2016, Detecting rumors from microblogs with recurrent neural networks, 3818 Bondielli, 2019, A survey on fake news and rumour detection techniques, Inform. Sci., 497, 38, 10.1016/j.ins.2019.05.035 Rezaee, 2020, A discrete variational recurrent topic model without the reparametrization trick, Adv. Neural Inf. Process. Syst., 33, 13831 Wang, 2020, Neural topic model with attention for supervised learning, 1147 Yang, 2022, An effective dimensionality reduction approach for short-term load forecasting, Electr. Power Syst. Res., 210, 10.1016/j.epsr.2022.108150 Dib, 2021, Incorporating LDA with LSTM for followee recommendation on Twitter network, Int. J. Web Inf. Syst., 10.1108/IJWIS-12-2020-0079 Jo, 2017 Hoffman, 2013, Stochastic variational inference, J. Mach. Learn. Res., 14 Chawla, 2002, SMOTE: Synthetic minority over-sampling technique, J. Artificial Intelligence Res., 16, 321, 10.1613/jair.953 K. Zhao, Z. Xu, M. Yan, Y. Tang, M. Fan, G. Catolino, Just-in-time defect prediction for Android apps via imbalanced deep learning model, in: Proceedings of the 36th Annual ACM Symposium on Applied Computing, 2021, pp. 1447–1454. Can, 2019, A new direction in social network analysis: Online social network analysis problems and applications, Phys. A Stat. Mech. Appl., 535, 10.1016/j.physa.2019.122372 Verdoliva, 2020, Media forensics and deepfakes: an overview, IEEE J. Sel. Top. Sign. Proces., 14, 910, 10.1109/JSTSP.2020.3002101 Rehurek, 2011, Gensim–python framework for vector space modelling, 2 M. Röder, A. Both, A. Hinneburg, Exploring the space of topic coherence measures, in: Proceedings of the Eighth ACM International Conference on Web Search and Data Mining, 2015, pp. 399–408. Torabi Asr, 2019, Big Data and quality data for fake news and misinformation detection, Big Data Soc., 6, 10.1177/2053951719843310 Böhm, 2020 Z. Wang, Z. Duan, H. Zhang, C. Wang, L. Tian, B. Chen, M. Zhou, Friendly topic assistant for transformer based abstractive summarization, in: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing, EMNLP, 2020, pp. 485–497. Brankovic, 2018, A distributed feature selection algorithm based on distance correlation with an application to microarrays, IEEE/ACM Trans. Comput. Biol. Bioinform., 16, 1802 Hosseini, 2018 Brambilla, 2021, Conversation graphs in online social media, 97 Brambilla, 2022, Graph-based conversation analysis in social media, Big Data Cogn. Comput., 6, 113, 10.3390/bdcc6040113 Brambilla, 2021, The role of social media in long-running live events: The case of the Big Four fashion weeks dataset, Data Brief, 35, 10.1016/j.dib.2021.106840 Javadian Sabet, 2021, A multi-perspective approach for analyzing long-running live events on social media. A case study on the “Big Four” international fashion weeks, Online Soc. Netw. Media, 24 de Souza, 2020, A systematic mapping on automatic classification of fake news in social media, Soc. Netw. Anal. Min., 10, 1, 10.1007/s13278-020-00659-2