ELYZA-japanese-Llama-2-13B-fast-instruct: benchmark results

Provider: Other. Access: Open.

Unified ELO 1395 ± 21, rank #2181 of 2928 rated models, from 14 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard v1 - HellaSwag81.82Normalized accuracy (%) (10-shot)47.2
Open LLM Leaderboard v1 - WinoGrande75.93Accuracy (%) (5-shot)45
Open LLM Leaderboard v1 - ARC Challenge57.51Normalized accuracy (%) (25-shot)37.7
pfgen-bench - Completion Mode - Truthfulness0.64Truthfulness Score36.4
pfgen-bench - Completion Mode - Helpfulness0.07Helpfulness Score34.7
pfgen-bench - Completion Mode - Fluency0.5Fluency Score34.5
Open LLM Leaderboard v1 - MMLU54.52Accuracy (%) (5-shot)34.1
pfgen-bench - Completion Mode - Score0.4pfgen Score (mean of three)33.9
Open LLM Leaderboard v1 - TruthfulQA MC243.82MC2 (%) (0-shot)24.9
Open LLM Leaderboard v1 - GSM8K2.73Accuracy (%) (5-shot)20.7
pfgen-bench - QA Mode - Helpfulness0.05Helpfulness Score17.9
pfgen-bench - QA Mode - Score0.33pfgen Score (mean of three)14.5

Interactive version: theaggregate.ai/model?slug=elyza-japanese-llama-2-13b-fast-instruct · How It Works · Data refreshed daily, snapshot 2026-09-23.