gemma-2B: benchmark results

Provider: Google. Released 2024-02-21. Access: Open.

Unified ELO 1310 ± 1, rank #1334 of 1392 rated models, from 170 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
EuroEval Finnish NLU - Scandisent FI89.49Sentiment classification Score (%)53.7
Open Japanese LLM - Mbpp Pylint Check32.73Score (%)51.3
Open Japanese LLM - Wiki NER SET F15.31Score (%)51.2
EuroEval Dutch NLU - DBRD86.49Sentiment classification Score (%)49.2
AI Energy Score (Text Generation)4Energy Score (1-5)47.5
EuroEval Faroese NLU - FoSent23.93Sentiment classification Score (%)45.6
EuroEval Italian NLU - Sentipolc1647.86Sentiment classification Score (%)45.5
URIAL-Bench - Math3.3Judge Score (0-10)44.4
EuroEval English NLU - SQuAD76.66Reading comprehension Score (%)38.3
EuroEval Danish NLU - Angry Tweets38.91Sentiment classification Score (%)36.6
EuroEval Faroese NLU - FoQA30.06Reading comprehension Score (%)36.5
EuroEval Spanish NLU - MLQA ES53.45Reading comprehension Score (%)35.9

Interactive version: theaggregate.ai/model?slug=gemma-2b · How It Works · Data refreshed daily, snapshot 2026-09-05.