gemma-3-4B-pt Base — benchmark results
Google's pretrained 4B Gemma 3 base checkpoint (March 2025), multimodal with a SigLIP vision encoder and 128K context. Provider: Google. Released 2025-03-12. Access: Open.
Unified ELO 1453 ± 9, rank #981 of 1776 rated models, from 519 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval Slovak NLU - CSFD Sentiment SK | 65.92 | Sentiment classification Score (%) | 99 |
| EuroEval Czech NLU - CSFD Sentiment | 68.79 | Sentiment classification Score (%) | 98.2 |
| Open Arabic LLM - Aratrust Offensive | 95.65 | Accuracy (%) | 97.2 |
| EuroEval Finnish NLU - Scandisent FI | 92.64 | Sentiment classification Score (%) | 93.8 |
| EuroEval Ukrainian NLU - Cross Domain UK Reviews | 62.56 | Sentiment classification Score (%) | 93.8 |
| EuroEval Swedish NLU - Swerec | 79.23 | Sentiment classification Score (%) | 92 |
| Ukrainian LLM - Long Flores UK RU | 29.65 | Score (%) | 91.7 |
| EuroEval Lithuanian NLU - Atsiliepimai | 47.92 | Sentiment classification Score (%) | 90.8 |
| EuroEval Icelandic NLU - NQII | 56.87 | Reading comprehension Score (%) | 89.8 |
| Open Arabic LLM - Arabic MMLU LAW (Professional) | 75.8 | Accuracy (%) | 89.2 |
| EuroEval Lithuanian NLU - MultiWikiQA LT | 68.69 | Reading comprehension Score (%) | 88.2 |
| IndoBias | 71.4 | Ideology and Religion IND (self-reported) | 88 |
Interactive version: theaggregate.ai/model?slug=gemma-3-4b-pt-base · How the rankings work · Data refreshed daily, snapshot 2026-07-22.