Apr 2026· Knowledge Discovery and Data Mining· 1 citation· 47 references
Computer Science
TL;DR
MemeBridge, a curated dataset centered on U.S.-originated memes, is introduced, designed to capture two complementary perspectives: how Chinese participants interpret these memes, and how U.S. participants anticipate how people from other cultures might misunderstand them.
Abstract
Communicating across cultures is inherently challenging, especially through culturally dense and ambiguous formats like memes. While people expect large language models (LLMs) hold promise for bridging such gaps, existing benchmark datasets often fail to capture the cultural context necessary for accurate interpretation. To address this, we introduce MemeBridge, a curated dataset centered on U.S.-originated memes, designed to capture two complementary perspectives: (1) how Chinese participants interpret these memes, and (2) how U.S. participants anticipate how people from other cultures might misunderstand them. Here, context refers to implicit cultural knowledge—background beliefs, norms, and shared assumptions that shape meme comprehension. The dataset was constructed via a multi-stage crowdsourcing pipeline with rigorous validation, including human agreement checks and GPT-based classification verification. Each meme is annotated with sentiment, emotion, cultural significance, and knowledge type, providing rich supervision for downstream tasks. Notably, we observe that the anticipated misunderstandings from U.S. participants are often inaccurate, highlighting the asymmetries in cultural understanding and the challenges of adopting perspectives beyond one's own. This bidirectional framing -- focusing on both expression and perception -- enables more nuanced benchmarking of cross-cultural comprehension. Our probing of multiple LLMs reveals that while models developed in different cultural contexts exhibit partial cross-cultural understanding, they often struggle with sophisticated interpretations. By contrast, fine-tuning with MemeBridge improves model performance, underscoring the value of culturally grounded resources for training and evaluating LLMs in globally diverse settings.
This work introduces MemeCULT-1K, a multilingual benchmark of 1,000 South Asian memes in Bengali, English, and Hindi, where each meme is paired with a cultural context note and three human-written explanations, along with a supplementary set of 54 Bengali regional dialect memes.
Tawsif Tashwar Dipto, Mehedi Ahamed, Radib Bin Kabir et al.· 0 citations
Online company review platforms have gained significant importance as they provide transparent insights into corporate culture and employee satisfaction. However, analyzing this feedback at scale remains challenging due to the linguistic complexity of workplace-specific discourse and the lack of high-quality, multi-dim...
Khanh-Long Ho-Vuong, Nhat-Huy Dang, Do Bao et al.· International Conference on...· 0 citations
Online misogyny increasingly appears in multimodal formats such as memes. Memes combine text and images to convey humor, sarcasm, and ideology. Misogynistic meaning is often implicit and culturally grounded. A meme interpreted as harmful in one cultural setting may be perceived differently in another. Based on the comp...
The Cross-Cultural Misogynistic Meme Detection Grand Challenge, CC-MMD 2026, addresses the problem of identifying misogynistic content in multimodal memes across culturally diverse annotation perspectives. Existing misogyny detection benchmarks have advanced multimodal content moderation, but most assume a single groun...
Rahul Ponnusamy, Bhuvaneswari Sivagnanam, Anshid K. A. Kizhakkeparambil et al.· Proceedings of the 28th Inte...· 5 citations
This paper presents M-NLE, a compact knowledge-augmented multitask model for jointly detecting offensive memes and generating natural language explanations, and suggests that knowledge-augmented explanation generation is a practical direction for more interpretable offensive meme detection.
Dibyanayan Bandyopadhyay, Baban Gain, Samrat Mukherjee et al.· IEEE Access· 0 citations
A new method, called CW-Net, translates the reasoning process of an autonomous vehicle’s AI system into understandable concepts that explain its behavior.
Short chart specifications are easy to write, but often produce uninspiring results. Flint is an open-source visualization language that offers a middle path, letting AI agents create expressive charts from compact, human-editable specifications. The post Flint: A visualization language for the AI era appeared first on Microsoft Research.
Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.