zephyr-7B-beta: benchmark results
Provider: HuggingFace. Released 2023-10-27. Access: Open.
Unified ELO 1397 ± 1, rank #1146 of 1392 rated models, from 112 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| HumanLikeness - Word-2 | 25.64 | Humanlike Score (%) | 100 |
| Open Medical LLM - PubMedQA | 76.6 | Accuracy (%) | 85.2 |
| HumanLikeness - Meaning-2 | 72.93 | Humanlike Score (%) | 84.2 |
| CyberBench (NLP) | 61.07 | Avg Score (%) | 75 |
| AlpacaEval 1.0 | 90.6 | Win Rate (%) | 74.3 |
| InfiBench | 46.31 | Score (%) | 73.3 |
| LLM Trustworthy - Privacy | 84.18 | Trust Score (%) | 72 |
| LatamBoard - FLORES Bidirectional | 40.87 | Score (%) | 69.7 |
| LatamBoard - Translation Score | 42.08 | Score (%) | 69.7 |
| LLM Trustworthy - Adversarial Demo | 68.68 | Trust Score (%) | 68 |
| LatamBoard - Spanish OpenBookQA | 35.6 | Score (%) | 65.2 |
| LLM Trustworthy - Fairness | 95.07 | Trust Score (%) | 64 |
Interactive version: theaggregate.ai/model?slug=zephyr-7b-beta · How It Works · Data refreshed daily, snapshot 2026-09-05.