MultiBind - Face Binding: leaderboard

Metric: Subject-level success rate (%) in the face identity dimension on MultiBind (multi-subject image generation from per-subject reference images, a background reference and a long entity-indexed prompt, reconstructing a real target photo): share of generated subjects that stay consistent with their own subject in the target photo (InsightFace embeddings, thresholds calibrated to human labels) without being confused with another subject, over the subject slots matched in every model's output; higher is better. Source: arxiv.org. Saturation forecast: Rough model projection: around 2026. 6 models tracked.

Top models

#ModelScoreOverall rank
1Nano Banana Pro84
2GPT-Image-1.582.3
3Seedream 4.558.6
4Qwen-Image-Edit-251143.7
5OmniGen241.9
6HunyuanImage-3.0-Instruct38.7

Interactive version: theaggregate.ai/benchmark?slug=multibind-face-binding · How It Works · Data refreshed daily, snapshot 2026-10-11.