Nova Lite — benchmark results
Amazon's low-cost multimodal Bedrock model (December 2024) taking text, images, and video in a 300K context. Provider: Amazon. Released 2024-12-03. Access: API.
Unified ELO 1495 ± 19, rank #818 of 1776 rated models, from 79 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| HELM NaturalQuestions (Open) | 81.52 | F1 (%) | 98.9 |
| Ko-AgentBench - L6 Efficient Tool Utilization | 31.33 | Efficiency Score (%) | 92.9 |
| HELM NarrativeQA | 76.81 | F1 (%) | 81.1 |
| HELM (Stanford) | 70.8 | Mean Win Rate (%) | 77.8 |
| HELM WMT 2014 | 20.43 | BLEU-4 (%) | 65 |
| LLM Stats (DROP) | 80.2 | Score (%) | 64.3 |
| LLM Stats (ChartQA) | 86.8 | Score (%) | 60.9 |
| AA Omniscience | -40.42 | Score | 53.6 |
| LLM Stats (TextVQA) | 80.2 | Score (%) | 50 |
| Ko-AgentBench - L3 Sequential Tool Reasoning | 85 | PSM (%) | 42.9 |
| Ko-AgentBench - L4 Parallel Tool Reasoning | 56.67 | Coverage (%) | 42.9 |
| HELM NaturalQuestions (Closed) | 35.24 | F1 (%) | 41.1 |
Interactive version: theaggregate.ai/model?slug=nova-lite · How the rankings work · Data refreshed daily, snapshot 2026-07-22.