MemeBridge - Explanation Choice: leaderboard
Metric: Multiple-choice accuracy (%; pick the meme explanation among it, the misunderstanding its U.S. contributor anticipated and a GPT-4-written distractor; 621 U.S.-originated memes with U.S. crowd labels, default prompt (no role-play), the same test given to Chinese participants). Source: arxiv.org. Saturation forecast: Around December 2026. 4 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | GPT-4o | 75.4 |
Interactive version: theaggregate.ai/benchmark?slug=memebridge-explanation-choice · How It Works · Data refreshed daily, snapshot 2026-09-26.