Fake News Detection on Social Media: A Data Mining Perspective
2017/08/07 by Kai Shu, Amy Sliva, Shu, Kai +7 · 57 citations
Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #H.2.8 #Misinformation and Its Impacts #Sentiment Analysis and Opinion Mining #Social and Information Networks (cs.SI) #Spam and Phishing Detection #cs.AI #cs.SI
paper · pdf · doi:10.48550/arxiv.1708.01967
ACM SIGKDD Explorations Newsletter, 2017
openalex publication_date 2017/08/07 · arxiv created 2017/09/03 · arxiv updated 2017/09/05 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
Abstract
Social media for news consumption is a double-edged sword. On the one hand, its low cost, easy access, and rapid dissemination of information lead people to seek out and consume news from social media. On the other hand, it enables the wide spread of "fake news", i.e., low quality news with intentionally false information. The extensive spread of fake news has the potential for extremely negative impacts on individuals and society. Therefore, fake news detection on social media has recently become an emerging research that is attracting tremendous attention. Fake news detection on social media presents unique characteristics and challenges that make existing detection algorithms from traditional news media ineffective or not applicable. First, fake news is intentionally written to mislead readers to believe false information, which makes it difficult and nontrivial to detect based on news content; therefore, we need to include auxiliary information, such as user social engagements on social media, to help make a determination. Second, exploiting this auxiliary information is challenging in and of itself as users' social engagements with fake news produce data that is big, incomplete, unstructured, and noisy. Because the issue of fake news detection on social media is both challenging and relevant, we conducted this survey to further facilitate research on the problem. In this survey, we present a comprehensive review of detecting fake news on social media, including fake news characterizations on psychology and social theories, existing algorithms from a data mining perspective, evaluation metrics and representative datasets. We also discuss related research areas, open problems, and future research directions for fake news detection on social media.
Citations
Cited by
- Agentic Multi-Persona Framework for Evidence-Aware Fake News Detection
- Decoding Fake Narratives in Spreading Hateful Stories: A Dual-Head RoBERTa Model with Multi-Task Learning
- The Impact of Data Characteristics on GNN Evaluation for Detecting Fake News
- From Veracity to Diffusion: Adressing Operational Challenges in Moving From Fake-News Detection to Information Disorders
- Pooling Attention: Evaluating Pretrained Transformer Embeddings for Deception Classification
- Learning to Control Misinformation: a Closed-loop Approach for Misinformation Mitigation over Social Networks
- SPOT: An Annotated French Corpus and Benchmark for Detecting Critical Interventions in Online Conversations
- Can MLLMs Read the Room? A Multimodal Benchmark for Verifying Truthfulness in Multi-Party Social Interactions
- FakeZero: Real-Time, Privacy-Preserving Misinformation Detection for Facebook and X
- Misinformation Detection using Large Language Models with Explainability
- ADMIT: Few-shot Knowledge Poisoning Attacks on RAG-based Fact Checking
- Bridging the gap between marine science and policy: communicating for an informed society and decision-making
- A Survey on Automatic Credibility Assessment Using Textual Credibility Signals in the Era of Large Language Models
- The Hateful Memes Challenge: Detecting Hate Speech in Multimodal Memes
- DRES: Fake news detection by dynamic representation and ensemble selection
- Beyond Artificial Misalignment: Detecting and Grounding Semantic-Coordinated Multimodal Manipulations
- A Dynamic Knowledge Update-Driven Model with Large Language Models for Fake News Detection
- PolyTruth: Multilingual Disinformation Detection using Transformer-Based Language Models
- Are LLMs Enough for Hyperpartisan, Fake, Polarized and Harmful Content Detection? Evaluating In-Context Learning vs. Fine-Tuning
- Designing Effective AI Explanations for Misinformation Detection: A Comparative Study of Content, Social, and Combined Explanations
- HSFN: Hierarchical Selection for Fake News Detection building Heterogeneous Ensemble
- Towards a general diffusion-based information quality assessment model
- Prompt-Induced Linguistic Fingerprints for LLM-Generated Fake News Detection
- MM-COVID: A Multilingual and Multimodal Data Repository for Combating COVID-19 Disinformation
- Mining the Social Fabric: Unveiling Communities for Fake News Detection in Short Videos
- Enhancing Rumor Detection Methods with Propagation Structure Infused Language Model
- Graph-Augmented Language Model Framework for Health Misinformation Detection
- MM-FusionNet: Context-Aware Dynamic Fusion for Multi-modal Fake News Detection with Large Vision-Language Models
- Social Media Information Operations
- Agent-Based Exploration of Recommendation Systems in Misinformation Propagation
- ROBAD: Robust Adversary-aware Local-Global Attended Bad Actor Detection Sequential Model
- Principals editors científics als canals de Telegram : una aproximació a la detecció de canals falsos amb ChatGPT i DeepSeek
- Implicit and Indirect: Detecting Face-threatening and Paired Actions in Asynchronous Online Conversations
- Explainable AI for online disinformation detection: Insights from a design science research project
- A Data Set of Internet Claims and Comparison of their Sentiments with Credibility
- MisinfoTeleGraph: Network-driven Misinformation Detection for German Telegram Messages
- A Survey on False Information Detection: From A Perspective of Propagation on Social Networks
- Document-Level Tabular Numerical Cross-Checking: A Coarse-to-Fine Approach
- Detecting Sockpuppetry on Wikipedia Using Meta-Learning
- RoE-FND: A Case-Based Reasoning Approach with Dual Verification for Fake News Detection via LLMs
- ISMAF: Intrinsic-Social Modality Alignment and Fusion for Multimodal Rumor Detection
- Robustness Evaluation of Graph-based News Detection Using Network Structural Information
- KGAlign: Joint Semantic-Structural Knowledge Encoding for Multimodal Fake News Detection
- Spectral Analysis of Fake News Propagation
- False Information on Web and Social Media: A Survey
- The Alt-Right and Global Information Warfare
- Many Ways to Be Fake: Benchmarking Fake News Detection Under Strategy-Driven AI Generation
- Bridging Cognition and Emotion: Empathy-Driven Multimodal Misinformation Detection
- A joint learning framework for fake news detection
- RAGAT-Mind: A Multi-Granular Modeling Approach for Rumor Detection Based on MindSpore
- Large Language Models and Social Media Information Integrity: Opportunities, Challenges, and Research Directions
- Exploring Lightweight Interventions at Posting Time to Reduce the Sharing of Misinformation on Social Media
- Video-Bench: Human-Aligned Video Generation Benchmark
- TIFIN India at SemEval-2025: Harnessing Translation to Overcome Multilingual IR Challenges in Fact-Checked Claim Retrieval
- Fake News Detection with Different Models
- A Survey of Social Cybersecurity: Techniques for Attack Detection, Evaluations, Challenges, and Future Prospects
- RumorLens: Interactive Analysis and Validation of Suspected Rumors on Social Media
- Disinformation, Misinformation, and Fake News in Social Media
- Exploring Modality Disruption in Multimodal Fake News Detection
Related