zephyr-7B-alpha — benchmark results
Provider: HuggingFace. Released 2023-10-09. Access: Open.
Unified ELO 1362 ± 15, rank #1370 of 1776 rated models, from 49 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| HumanLikeness - Discourse-2 | 74.58 | Humanlike Score (%) | 89.5 |
| HumanLikeness - Meaning-2 | 73.13 | Humanlike Score (%) | 89.5 |
| HumanLikeness - Syntax-1 | 84.56 | Humanlike Score (%) | 89.5 |
| HumanLikeness - Word-2 | 22.88 | Humanlike Score (%) | 73.7 |
| RewardBench | 73.92 | Score (%) | 65.7 |
| HumanLikeness - Sound-2 | 61.83 | Humanlike Score (%) | 63.2 |
| Open LLM Leaderboard - IFEval | 51.91 | Score | 61.9 |
| HumanLikeness - Discourse-1 | 76.23 | Humanlike Score (%) | 57.9 |
| AlpacaEval 1.0 | 85.76 | Win Rate (%) | 56.4 |
| Open LLM Leaderboard - GPQA | 6.38 | Score | 53.5 |
| EconLogicQA | 0.23 | Accuracy | 52.9 |
| RewardBench Chat Hard | 62.5 | Accuracy (%) | 52.8 |
Interactive version: theaggregate.ai/model?slug=zephyr-7b-alpha · How the rankings work · Data refreshed daily, snapshot 2026-07-22.