PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
November 7, 2025Science Advances21 citationsOpen Access

Source framing triggers systematic bias in large language models

View Full Paper
FGFederico GermaniGSGiovanni Spitale

Key Points

Key points are not available for this paper at this time.

Abstract

Large language models (LLMs) are increasingly used to evaluate text, raising urgent questions about whether their judgments are consistent, unbiased, and robust to framing effects. Here, we examine inter- and intramodel agreement across four state-of-the-art LLMs tasked with evaluating 4800 narrative statements on 24 different topics of social, political, and public health relevance, for a total of 192,000 assessments. We manipulate the disclosed source of each statement to assess how attribution to either another LLM or a human author of specified nationality affects evaluation outcomes. Different LLMs display a remarkably high degree of inter- and intramodel agreement across topics, but this alignment breaks down when source framing is introduced. Attributing statements to Chinese individuals systematically lowers agreement scores across all models and, in particular, for DeepSeek Reasoner. Our findings show that LLMs' own judgment of agreement with narrative statements exhibit systematic bias from framing effects, with substantial implications for the neutrality and fairness of LLM-mediated information systems.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Germani et al. (2025) studied this question.

synapsesocial.com/papers/6a067883562f02339273aeeehttps://doi.org/10.1126/sciadv.adz2924
Ask AI
Helpful
Bookmark
Share
View Full Paper