Skip to content

MemeBridge: A Dataset for Benchmarking and Mitigating the Bidirectional Cultural Gap in Meme Interpretation

Apr 2026 · Knowledge Discovery and Data Mining · 1 citation · 47 references
Computer Science

TL;DR

MemeBridge, a curated dataset centered on U.S.-originated memes, is introduced, designed to capture two complementary perspectives: how Chinese participants interpret these memes, and how U.S. participants anticipate how people from other cultures might misunderstand them.

Abstract

Communicating across cultures is inherently challenging, especially through culturally dense and ambiguous formats like memes. While people expect large language models (LLMs) hold promise for bridging such gaps, existing benchmark datasets often fail to capture the cultural context necessary for accurate interpretation. To address this, we introduce MemeBridge, a curated dataset centered on U.S.-originated memes, designed to capture two complementary perspectives: (1) how Chinese participants interpret these memes, and (2) how U.S. participants anticipate how people from other cultures might misunderstand them. Here, context refers to implicit cultural knowledge—background beliefs, norms, and shared assumptions that shape meme comprehension. The dataset was constructed via a multi-stage crowdsourcing pipeline with rigorous validation, including human agreement checks and GPT-based classification verification. Each meme is annotated with sentiment, emotion, cultural significance, and knowledge type, providing rich supervision for downstream tasks. Notably, we observe that the anticipated misunderstandings from U.S. participants are often inaccurate, highlighting the asymmetries in cultural understanding and the challenges of adopting perspectives beyond one's own. This bidirectional framing -- focusing on both expression and perception -- enables more nuanced benchmarking of cross-cultural comprehension. Our probing of multiple LLMs reveals that while models developed in different cultural contexts exhibit partial cross-cultural understanding, they often struggle with sophisticated interpretations. By contrast, fine-tuning with MemeBridge improves model performance, underscoring the value of culturally grounded resources for training and evaluating LLMs in globally diverse settings.

Read PDF

Similar papers

#computer vision Preprint Sep 2026

MemeCULT-1K: Benchmarking South Asian Cultural Context and Humor Understanding of Multimodal Models

This work introduces MemeCULT-1K, a multilingual benchmark of 1,000 South Asian memes in Bengali, English, and Hindi, where each meme is paired with a cultural context note and three human-written explanations, along with a supplementary set of 54 Bengali regional dialect memes.

Tawsif Tashwar Dipto, Mehedi Ahamed, Radib Bin Kabir et al. · 0 citations
Conference Aug 2026

ViCorpReviews: A Benchmark Dataset for Multi-Dimensional Sentiment and Hate Speech Detection in Vietnamese Workplace Context

Online company review platforms have gained significant importance as they provide transparent insights into corporate culture and employee satisfaction. However, analyzing this feedback at scale remains challenging due to the linguistic complexity of workplace-specific discourse and the lack of high-quality, multi-dim...

Khanh-Long Ho-Vuong, Nhat-Huy Dang, Do Bao et al. · 0 citations
Open access Sep 2026

Classify and Label the Content from the Cross-Cultural Misogynistic Meme Detection Using Zero-shot Prompts

Online misogyny increasingly appears in multimodal formats such as memes. Memes combine text and images to convey humor, sarcasm, and ideology. Misogynistic meaning is often implicit and culturally grounded. A meme interpreted as harmful in one cultural setting may be perceived differently in another. Based on the comp...

Kong-Qiang Wang, Qing Tan, Peng Zhang · 0 citations
#large language models Book Open access Oct 2026

Overview of the CC-MMD 2026 Grand Challenge: Cross-Cultural Misogynistic Meme Detection in Multimodal Memes

The Cross-Cultural Misogynistic Meme Detection Grand Challenge, CC-MMD 2026, addresses the problem of identifying misogynistic content in multimodal memes across culturally diverse annotation perspectives. Existing misogyny detection benchmarks have advanced multimodal content moderation, but most assume a single groun...

Rahul Ponnusamy, Bhuvaneswari Sivagnanam, Anshid K. A. Kizhakkeparambil et al. · 5 citations
Review Open access 2026

M-NLE: Knowledge-Augmented Multitask Learning for Offensive Meme Detection and Explanation Generation

This paper presents M-NLE, a compact knowledge-augmented multitask model for jointly detecting offensive memes and generating natural language explanations, and suggests that knowledge-augmented explanation generation is a practical direction for more interpretable offensive meme detection.

Dibyanayan Bandyopadhyay, Baban Gain, Samrat Mukherjee et al. · 0 citations

Related blog posts

Microsoft Research Blog Jul 8, 2026

Flint: A visualization language for the AI era

Short chart specifications are easy to write, but often produce uninspiring results. Flint is an open-source visualization language that offers a middle path, letting AI agents create expressive charts from compact, human-editable specifications. The post Flint: A visualization language for the AI era appeared first on Microsoft Research.

MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.