MemeBridge - Explanation Choice: leaderboard

Metric: Multiple-choice accuracy (%; pick the meme explanation among it, the misunderstanding its U.S. contributor anticipated and a GPT-4-written distractor; 621 U.S.-originated memes with U.S. crowd labels, default prompt (no role-play), the same test given to Chinese participants). Source: arxiv.org. Saturation forecast: Around December 2026. 4 models tracked.

Top models

#ModelScore
1GPT-4o75.4

Interactive version: theaggregate.ai/benchmark?slug=memebridge-explanation-choice · How It Works · Data refreshed daily, snapshot 2026-09-26.