zephyr-7B-beta — benchmark results
Provider: HuggingFace. Released 2023-10-27. Access: Open.
Unified ELO 1334 ± 17, rank #1462 of 1776 rated models, from 108 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| HumanLikeness - Word-2 | 25.64 | Humanlike Score (%) | 100 |
| Open Medical LLM - PubMedQA | 76.6 | Accuracy (%) | 85.2 |
| HumanLikeness - Meaning-2 | 72.93 | Humanlike Score (%) | 84.2 |
| CyberBench (NLP) | 61.07 | Avg Score (%) | 75 |
| AlpacaEval 1.0 | 90.6 | Win Rate (%) | 74.3 |
| InfiBench | 46.31 | Score (%) | 73.3 |
| LLM Trustworthy - Privacy | 84.18 | Trust Score (%) | 72 |
| LatamBoard - FLORES Bidirectional | 40.87 | Score (%) | 69.7 |
| LatamBoard - Translation Score | 42.08 | Score (%) | 69.7 |
| LLM Trustworthy - Adversarial Demo | 68.68 | Trust Score (%) | 68 |
| Open Korean LLM Leaderboard | 294.93 | Average Score (%) | 66.4 |
| LatamBoard - Spanish OpenBookQA | 35.6 | Score (%) | 65.2 |
Interactive version: theaggregate.ai/model?slug=zephyr-7b-beta · How the rankings work · Data refreshed daily, snapshot 2026-07-22.