PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
July 4, 20241 citationsOpen Access

Seeing Like an AI: How LLMs Apply (and Misapply) Wikipedia Neutrality Norms

View Full Paper
JAJoshua AshkinazeRGRuijia GuanLKLaura Kurek

Key Points

Key points are not available for this paper at this time.

Abstract

Large language models (LLMs) are trained on broad corpora and then used in communities with specialized norms. Is providing LLMs with community rules enough for models to follow these norms? We evaluate LLMs' capacity to detect (Task 1) and correct (Task 2) biased Wikipedia edits according to Wikipedia's Neutral Point of View (NPOV) policy. LLMs struggled with bias detection, achieving only 64% accuracy on a balanced dataset. Models exhibited contrasting biases (some under- and others over-predicted bias), suggesting distinct priors about neutrality. LLMs performed better at generation, removing 79% of words removed by Wikipedia editors. However, LLMs made additional changes beyond Wikipedia editors' simpler neutralizations, resulting in high-recall but low-precision editing. Interestingly, crowdworkers rated AI rewrites as more neutral (70%) and fluent (61%) than Wikipedia-editor rewrites. Qualitative analysis found LLMs sometimes applied NPOV more comprehensively than Wikipedia editors but often made extraneous non-NPOV-related changes (such as grammar). LLMs may apply rules in ways that resonate with the public but diverge from community experts. While potentially effective for generation, LLMs may reduce editor agency and increase moderation workload (e.g., verifying additions). Even when rules are easy to articulate, having LLMs apply them like community members may still be difficult.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Ashkinaze et al. (2024) studied this question.

synapsesocial.com/papers/68e615e9b6db6435875a8f03https://doi.org/10.48550/arxiv.2407.04183
Ask AI
Helpful
Bookmark
Share
View Full Paper