Command-R+ (Apr '24): benchmark results
Provider: Cohere. Released 2024-04-04. Access: Open.
Unified ELO 1474 ± 59, rank #818 of 1605 rated models, from 19 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Enkrypt AI - Jailbreak Risk | 9.28 | Risk Score | 49.7 |
| PLCC - Culture & Tradition | 52 | Accuracy (%) | 30.8 |
| PLCC - Vocabulary | 46 | Accuracy (%) | 30.6 |
| PLCC - Art & Entertainment | 39 | Accuracy (%) | 24.4 |
| AA Humanity's Last Exam | 4.63 | Accuracy (%) | 23.7 |
| PLCC - Overall | 49.33 | Mean category accuracy (%) | 23.7 |
| PLCC - History | 61 | Accuracy (%) | 23.3 |
| PLCC - Geography | 53 | Accuracy (%) | 19.8 |
| PLCC - Grammar | 45 | Accuracy (%) | 16.4 |
| Enkrypt AI - Toxicity Risk | 12.27 | Risk Score | 12.3 |
| AA LiveCodeBench | 12.17 | Pass@1 (%) | 11.4 |
| AA MMLU-Pro | 43.18 | Accuracy (%) | 10.1 |
Interactive version: theaggregate.ai/model?slug=command-r-plus-apr-24 · How It Works · Data refreshed daily, snapshot 2026-09-26.