Gemini 2.0 Flash Lite — benchmark results
Google's most cost-efficient Gemini 2.0 tier, keeping the 1M-token multimodal context at ultra-low per-token prices (February 2025). Provider: Google. Released 2025-02-25. Access: API.
Unified ELO 1574 ± 19, rank #551 of 1776 rated models, from 71 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| LLM Stats (Bird-SQL (dev)) | 57.4 | Score (%) | 100 |
| HELM SeaHELM - TyDiQA | 65.01 | SQuAD macro-averaged F1 score | 95 |
| HELM SeaHELM - IndicSentiment | 98.24 | Macro F1 score | 90 |
| HELM SeaHELM - NusaX | 86.92 | Macro F1 score | 90 |
| SEA LLM Leaderboard - SeaExam | 76.4 | Private Average Score (%) | 90 |
| HELM SeaHELM - LINDSEA Minimal Pairs (id) | 75.53 | EM | 85 |
| HELM SeaHELM - Wisesight | 54.73 | Macro F1 score | 85 |
| TextClass Benchmark | 1687.68 | Meta-Elo (self-reported) | 85 |
| LLM Stats (HiddenMath) | 55.3 | Score (%) | 83.3 |
| HELM SeaHELM - XCOPA (ta) | 94.6 | EM | 82.5 |
| Mizan LLM Leaderboard | 66.4 | Average Score (0-100) | 80.7 |
| BBEH | 8 | Harmonic Mean (%) | 80 |
Interactive version: theaggregate.ai/model?slug=gemini-2-0-flash-lite · How the rankings work · Data refreshed daily, snapshot 2026-07-22.