GPT-4.5 — benchmark results
OpenAI's large GPT-4.5 non-reasoning model (February 2025), later retired from the API. Provider: OpenAI. Released 2025-02-27. Access: API.
Unified ELO 1586 ± 8, rank #510 of 1776 rated models, from 201 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| ProLLM - SQL Disambiguation | 53.1 | Score (%) | 100 |
| OpenVLM MMMU - Management | 83.3 | Accuracy (%) | 99.6 |
| OpenVLM MMBench V1.1 CN - Nature Relation | 96.7 | Accuracy (%) | 99.5 |
| OpenVLM MMMU - Clinical Medicine | 80 | Accuracy (%) | 99.5 |
| OpenVLM MMBench V1.1 CN - Spatial Relationship | 90.7 | Accuracy (%) | 99.3 |
| OpenVLM MMMU - Art and Design | 77.5 | Accuracy (%) | 99.1 |
| OpenVLM MMBench V1.1 EN - Spatial Relationship | 89.3 | Accuracy (%) | 98.9 |
| OpenVLM MMMU - Business | 84 | Accuracy (%) | 98.9 |
| OpenVLM AI2D - Rock Strata | 92.7 | Accuracy (%) | 98.6 |
| OpenVLM MMBench V1.1 CN - Image Emotion | 87.8 | Accuracy (%) | 98.6 |
| OpenVLM MMMU - Art | 86.7 | Accuracy (%) | 98.2 |
| OpenVLM MMMU - Economics | 86.7 | Accuracy (%) | 98.1 |
Interactive version: theaggregate.ai/model?slug=gpt-4-5 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.