Claude Opus 4.5 (20251101) — benchmark results
November 1, 2025 Claude Opus 4.5 snapshot, tracked when sources report the dated API model. Provider: Anthropic. Released 2025-11-01. Access: API.
Unified ELO 1767 ± 10, rank #170 of 1776 rated models, from 55 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Gorilla API Bench (BFCL) | 77.47 | Overall Accuracy (%) | 100 |
| Vals AI MGSM | 94.76 | Accuracy (%) | 98.3 |
| AraGen | 80.29 | 3C3H Score (%) | 94 |
| SEAL - SciPredict | 23.05 | Score | 92.9 |
| Icelandic LLM - WinoGrande-IS | 94.67 | Score (%) | 92.3 |
| Chatbot Arena (Text) | 1469 | Elo | 92.2 |
| WebApp1K Duo | 76.3 | Pass@1 (%) | 90.6 |
| Vals AI MortgageTax | 68.68 | Accuracy (%) | 89 |
| EQ-Bench Creative Writing v3 | 1598.2 | Elo | 88.5 |
| Icelandic LLM - Belebele-IS | 93.78 | Score (%) | 88.5 |
| EQ-Bench Longform Writing | 73.1 | Writing Score (0-100) | 85.7 |
| Icelandic LLM - GED | 74.5 | Score (%) | 85.7 |
Interactive version: theaggregate.ai/model?slug=claude-opus-4-5-20251101 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.