Command R — benchmark results

Cohere Command R model optimized for retrieval-augmented generation and enterprise chat. Provider: Cohere. Released 2024-03-11. Access: Open.

Unified ELO 1383 ± 16, rank #1300 of 1776 rated models, from 52 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LiveBench Zebra Puzzle38Score79.2
RULER88.9Avg Accuracy (%)61.3
CanAiCode98.9Junior-v2 Python Pass Rate (%)60.2
HELM NaturalQuestions (Open)72.05F1 (%)53.3
LiveBench Table Join9.08Score51.4
HELM NarrativeQA74.17F1 (%)50
IFEval Leaderboard69.71Final Score50
LiveBench Cta48Score50
LiveBench Typos22Score48.6
LiveBench Story Generation63.5Score47.2
MixEval45.2Score47.1
LiveBench Paraphrase59.85Score45.8

Interactive version: theaggregate.ai/model?slug=command-r · How the rankings work · Data refreshed daily, snapshot 2026-07-22.