zephyr-7B-beta: benchmark results

Provider: HuggingFace. Released 2023-10-27. Access: Open.

Unified ELO 1397 ± 1, rank #1146 of 1392 rated models, from 112 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
HumanLikeness - Word-225.64Humanlike Score (%)100
Open Medical LLM - PubMedQA76.6Accuracy (%)85.2
HumanLikeness - Meaning-272.93Humanlike Score (%)84.2
CyberBench (NLP)61.07Avg Score (%)75
AlpacaEval 1.090.6Win Rate (%)74.3
InfiBench46.31Score (%)73.3
LLM Trustworthy - Privacy84.18Trust Score (%)72
LatamBoard - FLORES Bidirectional40.87Score (%)69.7
LatamBoard - Translation Score42.08Score (%)69.7
LLM Trustworthy - Adversarial Demo68.68Trust Score (%)68
LatamBoard - Spanish OpenBookQA35.6Score (%)65.2
LLM Trustworthy - Fairness95.07Trust Score (%)64

Interactive version: theaggregate.ai/model?slug=zephyr-7b-beta · How It Works · Data refreshed daily, snapshot 2026-09-05.