Phi-3 Mini Instruct 3.8B: benchmark results
Provider: Microsoft. Released 2024-04-23. Access: Open.
Unified ELO 1372 ± 1, rank #1614 of 1919 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA Humanity's Last Exam | 4.96 | Accuracy (%) | 29.5 |
| Artificial Analysis Intelligence Index | 5.82 | Intelligence Index | 13.5 |
| AA GPQA Diamond | 31.92 | Accuracy (%) | 7.8 |
| AA IFBench | 23.88 | Accuracy (%) | 6.2 |
| AA Terminal-Bench Hard | 0 | Accuracy (%) | 5.5 |
| BenchmarkList ECI | 68.7 | Capability Index (ECI) | 4.3 |
| AA TAU-2 Bench | 0 | Accuracy (%) | 2.5 |
Interactive version: theaggregate.ai/model?slug=phi-3-mini-instruct-3-8b · How It Works · Data refreshed daily, snapshot 2026-09-08.