Mistral Medium 3 — benchmark results

Mistral's enterprise multimodal mid-tier positioned near Claude 3.7 Sonnet performance at much lower cost, deployable on-prem (May 2025). Provider: Mistral. Released 2025-05-07. Access: API.

Unified ELO 1553 ± 12, rank #612 of 1776 rated models, from 121 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
BlueBench - Legal67.17Score (%)100
UGI Leaderboard56.45UGI Score97.1
BlueBench - QA Finance33Score (%)91.2
UGI - Natural Intelligence35.26NatInt Score84
BlueBench - Chatbot Abilities94.17Score (%)82.4
BlueBench - Entity Extraction74.21Score (%)82.4
BlueBench - RAG General52.57Score (%)82.4
BlueBench - Reasoning75.5Score (%)79.4
BlueBench60.5Average Score (%)76.5
YapBench361YapIndex (lower is better)73.8
UGI - Willingness (W/10)7.2W/10 Score73.6
AA MATH-50090.67Accuracy (%)66.8

Interactive version: theaggregate.ai/model?slug=mistral-medium-3 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.