Gemma 3 4B — benchmark results
Google Gemma 3 4B model row. Provider: Google. Released 2025-03-12. Access: Open.
Unified ELO 1394 ± 6, rank #1254 of 1776 rated models, from 567 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| TinyQABenchmark (core-en) | 86.5 | Exact Match | 100 |
| OpenVLM LLaVA-Bench - Detail | 96.6 | Score | 98.5 |
| OpenVLM MMT-Bench - Plant Recognition | 95 | Score (%) | 98.1 |
| Open LMM Reasoning - WeMath - InadequateGeneralization (Loose) | 15.6 | Accuracy (%) | 95.7 |
| OpenVLM MMT-Bench - Shape Recognition | 95 | Score (%) | 93.7 |
| OpenVLM MMT-Bench - Font Recognition | 40 | Score (%) | 93.2 |
| OpenVLM MMT-Bench - Tech Engineering | 50 | Score (%) | 93 |
| OpenVLM MMT-Bench - Temporal Anticipation | 80 | Score (%) | 91.3 |
| OpenVLM MMBench V1.1 CN - Image Topic | 98.9 | Accuracy (%) | 90.3 |
| OpenVLM MMT-Bench - LVLM Response Judgement | 60 | Score (%) | 88.6 |
| OpenVLM LLaVA-Bench - Complex | 100 | Score | 88.4 |
| Open LMM Reasoning - WeMath - Calculation of Solid Figures | 81.7 | Accuracy (%) | 87.7 |
Interactive version: theaggregate.ai/model?slug=gemma-3-4b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.