PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 24, 2024200 citationsOpen Access

Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection

View Full Paper
BHBeizhe HuQSQiang ShengJCJuan Cao

Key Points

  • Large language models generate helpful multi-perspective rationales but fail to outperform fine-tuned small language models when classifying fake news directly.
  • Benchmark on two real-world datasets shows the adaptive rationale guidance network and its distilled variant outperform standalone language models.
  • Structuring large language models as advisory systems allows small language models to integrate multi-perspective rationales with reduced computational cost.

Abstract

Detecting fake news requires both a delicate sense of diverse clues and a profound understanding of the real-world background, which remains challenging for detectors based on small language models (SLMs) due to their knowledge and capability limitations. Recent advances in large language models (LLMs) have shown remarkable performance in various tasks, but whether and how LLMs could help with fake news detection remains underexplored. In this paper, we investigate the potential of LLMs in fake news detection. First, we conduct an empirical study and find that a sophisticated LLM such as GPT 3.5 could generally expose fake news and provide desirable multi-perspective rationales but still underperforms the basic SLM, fine-tuned BERT. Our subsequent analysis attributes such a gap to the LLM's inability to select and integrate rationales properly to conclude. Based on these findings, we propose that current LLMs may not substitute fine-tuned SLMs in fake news detection but can be a good advisor for SLMs by providing multi-perspective instructive rationales. To instantiate this proposal, we design an adaptive rationale guidance network for fake news detection (ARG), in which SLMs selectively acquire insights on news analysis from the LLMs' rationales. We further derive a rationale-free version of ARG by distillation, namely ARG-D, which services cost-sensitive scenarios without inquiring LLMs. Experiments on two real-world datasets demonstrate that ARG and ARG-D outperform three types of baseline methods, including SLM-based, LLM-based, and combinations of small and large language models.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Hu et al. (2024) studied this question.

synapsesocial.com/papers/68e72954b6db6435876a2cc1https://doi.org/10.1609/aaai.v38i20.30214
Ask AI
Helpful
Bookmark
Share
View Full Paper