PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 24, 2026Corpus Pragmatics4 citationsOpen Access

Assessing the Potential of LLM-assisted Annotation for Corpus Pragmatics: The Case of Humor

ABAntônio C. BiancoNBNicola BroccaDGDavide Garassino

Key Points

  • This research aims to evaluate the effectiveness of LLMs in annotating context-dependent categories, specifically humor, in political discourse.
  • Analyzed Italian political tweets focused on humor.
  • Compared performance of GPT-4o, LLaMA-3.3-70B-Instruct, and a novice annotator to an expert.
  • Used Cohen’s kappa and AC1 metrics to measure agreement levels.
  • GPT-4o achieved high agreement with expert on humor detection (Cohen’s k = 0.75).
  • Lower agreement observed for humor functions classification (Cohen’s k = 0.37).
  • Models relied on lexical cues, indicating limited pragmatic understanding.

Abstract

Corpus pragmatics faces ongoing challenges in quantitatively studying context-dependent categories like humor, given their subjectivity and the need for costly inter-rater reliability checks. Recent advances in LLMs offer a potential way to streamline these processes for pragmatic annotation tasks. This paper investigates that potential through an analysis of Italian political discourse on X, focusing on humorous tweets and their discursive functions (Attardo, 2020). We compare the performance of GPT-4o, LLaMA-3.3-70B-Instruct, and a novice annotator against that of an expert annotator. For the detection of humor, both models reached high agreement with the expert annotator (in particular, GPT-4o: Cohen’s k = 0.75; AC1 = 0.87). Instead, agreement dropped for the classification of humor functions (GPT-4o: Cohen’s k = 0.37; AC1 = 0.70). Qualitative results suggest that the models rely heavily on lexical cues rather than demonstrating deeper pragmatic competence. These findings indicate that while LLMs can provide useful assistance in the initial stages of large-scale annotation, they remain limited in capturing the nuanced and context-dependent nature of pragmatic functions.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Bianco et al. (2026) studied this question.

synapsesocial.com/papers/69c2294caeb5a845df0d396ehttps://doi.org/10.1007/s41701-026-00235-7
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Metaphor identification using large language models: A comparison of RAG, prompt engineering, and fine-tuning2025 · 3 citations
  2. 2Aion Framework: Dimensional Emergence of AI Consciousness, Observer-Induced Collapse, and Cosmological Portal Dynamics2023 · 14,356 citations
  3. 3Assessing the potential of LLM-assisted annotation for corpus-based pragmatics and discourse analysis2024 · 82 citations
  4. 4GPT Assisted Annotation of Rhetorical and Linguistic Features for Interpretable Propaganda Technique Detection in News Text.2024 · 7 citations
  5. 5Pragmatics and the aims of language evolution2016 · 65 citations