Synapse
⌘+K
Synapse
PulseExploreClubsResearchersJournals
Instagram
HomeClubsExplore
October 20, 2025Open Access

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia

View Full Paper
Ask AI
Bookmark
Share

Authors

DCDaniel P. CostaNOAA National Marine Fisheries ServiceRVRafael J. VicenteUniversidad Politécnica de Madrid

Discussion

Loading...

Member takes

Implication

This analysis demonstrates the interactive capabilities of large language models in detecting deception in Mini-Mafia, indicating important implications for AI safety.

Key Points

  • Mini-Mafia enables large language models to interactively demonstrate deception and detection skills, enhancing our understanding of their social intelligence.
  • Experimental findings show that smaller models can outperform larger ones, challenging traditional assumptions about model size and capability in deception tasks.
  • The Mini-Mafia Benchmark offers a systematic approach to evaluate the performance of language models across fixed opponent configurations, enhancing testing reliability.
  • This approach contributes to AI safety by generating valuable training data for deception detectors and evaluating model capabilities against human baselines.

Cite This Study

Costa et al. (2025) studied this question.

synapsesocial.com/papers/68f6196ee0bbbc94fac3647fhttps://doi.org/10.48550/arxiv.2509.23023
View Full Paper
Ask AI
Bookmark
Share

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Among LLMs: A Cross-Play Benchmark for Deception, Detection, and the Monitorability of Reasoning2026
  2. 2Among LLMs: A Cross-Play Benchmark for Deception, Detection, and the Monitorability of Reasoning2026
  3. 3Microscopic Analysis on LLM players via Social Deduction Game2024
  4. 4Training compact language models for artificial emotional intelligence: from bluffing to trust in a social deduction game2025
  5. 5Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?2024 · 1 citations