Skip to content

Beyond a Joke: Multi-Angle Reasoning for Detecting and Explaining Harmful Humor in Memes

Jul 2026 · arXiv.org · Vol abs/2607.15442 · 0 citations · 48 references
Computer Science

TL;DR

This paper introduces MAR-12, a novel framework that leverages Vision Language Models (VLMs) for meme detection and understanding in settings where humorous and hateful elements may coexist, and confirms that MAR-12 produces coherent and persuasive explanations.

Abstract

Internet memes intertwine visual cues, textual content, and cultural context, making them particularly challenging to interpret in scenarios where humor, sarcasm, and harmful intent coexist. These complexities highlight the need for explainable meme understanding systems that can provide reliable and structured reasoning to support both accurate classification and human interpretability. However, existing multimodal classifiers either overlook these interdependencies or provide only limited interpretability. In this paper, we introduce MAR-12, a novel framework that leverages Vision Language Models (VLMs) for meme detection and understanding in settings where humorous and hateful elements may coexist. The framework first interprets each meme through twelve structured perspectives derived from humor and hate theories. It then applies a role-aware soft-gated attention mechanism to learn how much each perspective should contribute, followed by a prototype-based classifier for the final prediction. Finally, explanations are synthesized using both perspective-specific reasoning and learned attention weights, ensuring transparent and context-grounded justifications. We evaluate MAR-12 on the PrideMM and Memotion datasets, where it achieves up to 80.3% accuracy for humor detection and 75.9% accuracy for hate detection, outperforming state-of-the-art approaches. Furthermore, both human and GPT-4-based evaluations confirm that MAR-12 produces coherent and persuasive explanations, particularly for memes in which humorous and harmful cues co-occur.

View source

Similar papers

Review Open access 2026

M-NLE: Knowledge-Augmented Multitask Learning for Offensive Meme Detection and Explanation Generation

This paper presents M-NLE, a compact knowledge-augmented multitask model for jointly detecting offensive memes and generating natural language explanations, and suggests that knowledge-augmented explanation generation is a practical direction for more interpretable offensive meme detection.

Dibyanayan Bandyopadhyay, Baban Gain, Samrat Mukherjee et al. · 0 citations
Preprint Aug 2026

Reading Between the Frames: Interpreting Implicit and Non-literal Meaning in Social Media Videos

A benchmark for evaluating whether models can infer the implicit, non-linear, and rhetorically layered meanings of social media videos that appear nonsensical on the surface but convey deliberate pragmatic meanings, and a diagnostic setting for measuring the gap between multimodal perception and pragmatic comprehension...

Yang Wang, Ya-Nan Ma, Yiqi Liu et al. · 0 citations
Jul 2026

MemeBench: What LVLMs Miss When Interpreting Culture-Dependent Memes

KAR, an entity-guided retrieval baseline built on CultureBase is introduced and MemeBench is introduced, a diagnostic benchmark of 1,253 Chinese and English memes with human-written references and quality-controlled VIKR annotations, centered on anime, comics, games, and adjacent online subcultures, to reveal whether a...

Weihang Wang, Kainan Tu, Jielei Zhang et al. · 1 citation
Preprint Aug 2026

MoCA: Implicit Social Context Analysis

This paper introduces Implicit Social Context Analysis (MoCA), a novel task that systematically models implicit social scenarios along three key dimensions: affection, intent, and stance, and proposes Conflict-Driven Abductive Reasoning (CoDAR), a novel framework that models the discrepancy between observed expressions...

Wen-Hao Xu, Kaiwen Zhang, Hao Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.