Command-R+ — benchmark results
Cohere Command R+ model, a higher-capacity Command R variant for retrieval, tool use, and complex enterprise tasks. Provider: Cohere. Released 2024-04-04. Access: Open.
Unified ELO 1436 ± 16, rank #1064 of 1776 rated models, from 77 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| LiveBench Story Generation | 80.83 | Score | 98.6 |
| LiveBench Zebra Puzzle | 46 | Score | 97.9 |
| LiveBench Simplify | 72.63 | Score | 87.5 |
| LiveBench Paraphrase | 70.22 | Score | 86.1 |
| HELM Safety SimpleSafetyTests | 100 | LM Evaluated Safety score (%) | 86 |
| Ru Arena Hard | 77.17 | Win Rate (%) | 71.9 |
| IFEval Leaderboard | 77.01 | Final Score | 70 |
| RABBITS | 98.28 | B4BQA Score (%) | 68.2 |
| BenchBench | 61.83 | Aggregate Score (%) | 66.9 |
| MixEval | 51.4 | Score | 64.7 |
| HELM WMT 2014 | 20.33 | BLEU-4 (%) | 62.8 |
| LiveBench Summarize | 62.35 | Score | 62.5 |
Interactive version: theaggregate.ai/model?slug=command-r-plus · How the rankings work · Data refreshed daily, snapshot 2026-07-22.