gemma-2B — benchmark results
Provider: Google. Released 2024-02-21. Access: Open.
Unified ELO 1249 ± 19, rank #1619 of 1776 rated models, from 171 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open Japanese LLM - ALT J TO E Bleu EN | 31.09 | Score (%) | 98.3 |
| Open Japanese LLM - Wikicorpus J TO E Bleu EN | 33.83 | Score (%) | 97.6 |
| Open Japanese LLM - Wikicorpus E TO J Bleu JA | 41.92 | Score (%) | 96.4 |
| Open Japanese LLM - ALT E TO J Bleu JA | 23.5 | Score (%) | 94.5 |
| EuroEval Finnish NLU - Scandisent FI | 89.49 | Sentiment classification Score (%) | 53.7 |
| Open Japanese LLM - Mbpp Pylint Check | 32.73 | Score (%) | 51.3 |
| Open Japanese LLM - Wiki NER SET F1 | 5.31 | Score (%) | 51.2 |
| EuroEval Dutch NLU - DBRD | 86.49 | Sentiment classification Score (%) | 49.2 |
| AI Energy Score (Text Generation) | 4 | Energy Score (1-5) | 47.5 |
| EuroEval Faroese NLU - FoSent | 23.93 | Sentiment classification Score (%) | 45.6 |
| EuroEval Italian NLU - Sentipolc16 | 47.86 | Sentiment classification Score (%) | 45.5 |
| URIAL-Bench - Math | 3.3 | Judge Score (0-10) | 44.4 |
Interactive version: theaggregate.ai/model?slug=gemma-2b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.