Nova Pro: benchmark results
Amazon's mid-tier Nova multimodal model handling text, image and video input with a 300K context on Bedrock (December 2024). Provider: Amazon. Released 2024-12-03. Access: API.
Unified ELO 1529 ± 1, rank #549 of 1392 rated models, from 102 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| HELM NaturalQuestions (Open) | 82.9 | F1 (%) | 100 |
| Translation en to Set1 spBleu | 43.4 | en→Set1 spBLEU (self-reported) | 100 |
| HELM (Stanford) | 88.48 | Mean Win Rate (%) | 95.6 |
| HELM NarrativeQA | 79.13 | F1 (%) | 95 |
| Translation Set1 to en spBleu | 44.4 | Set1→en spBLEU (self-reported) | 91.7 |
| HELM WMT 2014 | 22.89 | BLEU-4 (%) | 90 |
| VIBench | 12.2 | Direct VIB (self-reported) | 88.9 |
| LLM Stats (ChartQA) | 89.2 | Score (%) | 88 |
| ProLLM - Entity Extraction | 84.5 | Score (%) | 87.2 |
| LLM Stats (DROP) | 85.4 | Score (%) | 85.7 |
| Ukrainian Legal Text Benchmark (EDRSR subset) | 82.7 | 3-task Composite (self-reported) | 83.3 |
| Translation Set1 to en COMET22 | 89 | Set1→en COMET22 (self-reported) | 79.2 |
Interactive version: theaggregate.ai/model?slug=nova-pro · How It Works · Data refreshed daily, snapshot 2026-09-05.