PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 20, 2025Applied Sciences1 citationsOpen Access

Multimodal Chinese Sarcasm Detection Integrating Audio Attributes and Textual Features

View Full Paper
HWHuixin WuLZLimeng ZhaoHZHongyang Zhou

Key Points

Key points are not available for this paper at this time.

Abstract

Sarcasm often arises from subtle contrasts between literal meaning and speaker intention. As online communication increasingly includes voice-based content, detecting sarcasm across speech and text becomes more important—and more complex. The existing methods usually focus on generic multimodal fusion but often miss how sarcasm manifests differently in each modality. We propose a model that explicitly encodes audio signals into the textual representation space, allowing prosodic cues to inform language understanding. To extract relevant features at different levels, we use a multi-scale convolutional architecture. The experiments show consistent gains over prior models on both text and speech sarcasm detection tasks.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Wu et al. (2025) studied this question.

synapsesocial.com/papers/6a8646e50738f7dfc35ae664https://doi.org/10.3390/app15105689
Ask AI
Helpful
Bookmark
Share
View Full Paper