zephyr-7B-beta — benchmark results

Provider: HuggingFace. Released 2023-10-27. Access: Open.

Unified ELO 1334 ± 17, rank #1462 of 1776 rated models, from 108 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
HumanLikeness - Word-225.64Humanlike Score (%)100
Open Medical LLM - PubMedQA76.6Accuracy (%)85.2
HumanLikeness - Meaning-272.93Humanlike Score (%)84.2
CyberBench (NLP)61.07Avg Score (%)75
AlpacaEval 1.090.6Win Rate (%)74.3
InfiBench46.31Score (%)73.3
LLM Trustworthy - Privacy84.18Trust Score (%)72
LatamBoard - FLORES Bidirectional40.87Score (%)69.7
LatamBoard - Translation Score42.08Score (%)69.7
LLM Trustworthy - Adversarial Demo68.68Trust Score (%)68
Open Korean LLM Leaderboard294.93Average Score (%)66.4
LatamBoard - Spanish OpenBookQA35.6Score (%)65.2

Interactive version: theaggregate.ai/model?slug=zephyr-7b-beta · How the rankings work · Data refreshed daily, snapshot 2026-07-22.