Llama 3.2 Vision 11B Instruct: benchmark results
Provider: Meta. Released 2024-09-25. Access: Open.
Unified ELO 1414 ± 35, rank #2019 of 2928 rated models, from 13 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| BenchTable - Utility | 56.6 | Weighted Score (%) | 56.5 |
| SnakeBench | 18 | TrueSkill Rating | 31.5 |
| Wolfram LLM Benchmarking Project | 23.9 | Correct Functionality (%) | 21.8 |
| BenchTable - Tech | 19.7 | Weighted Score (%) | 21.5 |
| BenchTable | 18.6 | Total Score (%) | 14.4 |
| BenchmarkList ECI | 75.25 | Capability Index (ECI) | 8.3 |
| Chatbot Arena (Vision - English) | 1016 | Arena Score | 7.3 |
| Chatbot Arena (Vision) | 993 | Arena Score | 5.3 |
| BenchTable - Reasoning | -1.6 | Weighted Score (%) | 3.1 |
| BenchTable - STEM | 1.9 | Weighted Score (%) | 2.9 |
| SEAL - VISTA | 20.47 | Score | 1.6 |
| BLUEX v2 | 4.18 | Score (self-reported) | 0 |
Interactive version: theaggregate.ai/model?slug=llama-3-2-vision-11b-instruct · How It Works · Data refreshed daily, snapshot 2026-09-23.