MAR-12: New AI Model Detects and Explains Harmful Humor in Memes
Researchers introduced MAR-12, an AI model that analyzes memes to detect harmful humor and provides structured reasoning for its classifications. By combining visual, textual, and cultural context, MAR-12 aims to improve online safety and moderation with explainable AI.

Researchers have introduced MAR-12, a new AI model designed to detect and explain harmful humor in internet memes. Memes combine images, text, and cultural context, making them particularly challenging for AI to interpret. Existing multimodal classifiers often miss the nuances of sarcasm, humor, and harmful intent, or provide only limited interpretability. MAR-12 addresses this by offering structured reasoning to support both accurate classification and human understanding.
This matters because harmful memes can spread quickly online, causing emotional distress or reinforcing harmful stereotypes. An AI that not only flags offensive content but also explains its reasoning can help users understand the impact of their humor. This could lead to better online moderation and more thoughtful content creation.
For technical details, the full paper titled 'Beyond a Joke: Multi-Angle Reasoning for Detecting and Explaining Harmful Humor in Memes' is available on arXiv.