Command R — benchmark results
Cohere Command R model optimized for retrieval-augmented generation and enterprise chat. Provider: Cohere. Released 2024-03-11. Access: Open.
Unified ELO 1383 ± 16, rank #1300 of 1776 rated models, from 52 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| LiveBench Zebra Puzzle | 38 | Score | 79.2 |
| RULER | 88.9 | Avg Accuracy (%) | 61.3 |
| CanAiCode | 98.9 | Junior-v2 Python Pass Rate (%) | 60.2 |
| HELM NaturalQuestions (Open) | 72.05 | F1 (%) | 53.3 |
| LiveBench Table Join | 9.08 | Score | 51.4 |
| HELM NarrativeQA | 74.17 | F1 (%) | 50 |
| IFEval Leaderboard | 69.71 | Final Score | 50 |
| LiveBench Cta | 48 | Score | 50 |
| LiveBench Typos | 22 | Score | 48.6 |
| LiveBench Story Generation | 63.5 | Score | 47.2 |
| MixEval | 45.2 | Score | 47.1 |
| LiveBench Paraphrase | 59.85 | Score | 45.8 |
Interactive version: theaggregate.ai/model?slug=command-r · How the rankings work · Data refreshed daily, snapshot 2026-07-22.