c4ai-command-r-v01 — benchmark results
CohereForAI Command R v01 open-weight checkpoint. Provider: Cohere. Released 2024-03-11. Access: Open.
Unified ELO 1451 ± 17, rank #992 of 1776 rated models, from 63 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - MuSR | 16.13 | Score | 85 |
| Open LLM Leaderboard - IFEval | 67.48 | Score | 79.9 |
| BiGGen-Bench | 3.68 | Average Score (1-5) | 73.5 |
| EuroEval Portuguese NLU - HAREM | 48.64 | Named entity recognition Score (%) | 73.3 |
| EuroEval Dutch NLU - DBRD | 90.19 | Sentiment classification Score (%) | 72.6 |
| BABILong (NIAH) | 55.3 | Avg Accuracy (%) | 72.4 |
| EuroEval Italian NLU - MultiNERD IT | 69.33 | Named entity recognition Score (%) | 70.3 |
| EuroEval Spanish NLU - Sentiment Headlines ES | 45.06 | Sentiment classification Score (%) | 69.7 |
| Open LLM Leaderboard - BBH | 34.56 | Score | 68.7 |
| EuroEval Spanish NLU | 47.43 | NLU Average Score (%) | 66.3 |
| EuroEval Portuguese NLU | 53.05 | NLU Average Score (%) | 66.1 |
| Open PL LLM - Generative | 57.25 | Average Generative Score (%) | 64.7 |
Interactive version: theaggregate.ai/model?slug=c4ai-command-r-v01 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.