llm-jp-13B-instruct-full-jaster-v1.0: benchmark results

Provider: Other. Access: Open.

Unified ELO 1191 ± 20, rank #2913 of 2928 rated models, from 14 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
pfgen-bench - Completion Mode - Truthfulness0.8Truthfulness Score61.8
pfgen-bench - Completion Mode - Score0.5pfgen Score (mean of three)57.5
pfgen-bench - Completion Mode - Helpfulness0.13Helpfulness Score53
pfgen-bench - Completion Mode - Fluency0.58Fluency Score52.4
pfgen-bench - QA Mode - Truthfulness0.69Truthfulness Score37
Open LLM Leaderboard v1 - TruthfulQA MC244.69MC2 (%) (0-shot)28.1
pfgen-bench - QA Mode - Score0.28pfgen Score (mean of three)10.6
Open LLM Leaderboard v1 - HellaSwag44.7Normalized accuracy (%) (10-shot)9
pfgen-bench - QA Mode - Helpfulness0.01Helpfulness Score6.7
Open LLM Leaderboard v1 - ARC Challenge27.22Normalized accuracy (%) (25-shot)6.4
Open LLM Leaderboard v1 - GSM8K0Accuracy (%) (5-shot)4.6
Open LLM Leaderboard v1 - WinoGrande50.04Accuracy (%) (5-shot)2.8

Interactive version: theaggregate.ai/model?slug=llm-jp-13b-instruct-full-jaster-v1-0 · How It Works · Data refreshed daily, snapshot 2026-09-23.