Llama 3.2 90B — benchmark results
Meta's 90B vision-language base model from Llama 3.2, adding cross-attention image adapters to Llama text weights with a 128K context. Provider: Meta. Released 2024-09-25. Access: Open.
Unified ELO 1457 ± 27, rank #965 of 1776 rated models, from 21 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Translation Set1 to en spBleu | 43.7 | Set1âen spBLEU (self-reported) | 75 |
| AI for Education Pedagogy - Technology | 82.08 | Accuracy (%) | 71.6 |
| MMMU Benchmark | 60.3 | Validation Score | 66.7 |
| AI for Education Pedagogy - Maths | 76.19 | Accuracy (%) | 50.9 |
| AI for Education Pedagogy | 76.31 | Accuracy (%) | 49.1 |
| AI for Education Pedagogy - Secondary | 75 | Accuracy (%) | 47.5 |
| AI for Education Pedagogy - Social studies | 74.55 | Accuracy (%) | 47 |
| AI for Education SEND | 70.64 | Accuracy (%) | 44.6 |
| AI for Education Pedagogy - Primary | 79.81 | Accuracy (%) | 42.7 |
| AI for Education Pedagogy - Science | 74.86 | Accuracy (%) | 38.6 |
| AI for Education Visual Maths - Geometry | 36.92 | Accuracy (%) | 34.2 |
| AI for Education Visual Maths - Statistics and Probability | 14.29 | Accuracy (%) | 29.2 |
Interactive version: theaggregate.ai/model?slug=llama-3-2-90b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.