The AI arena is free today

Open Playground
CohereReleased on Aug 30, 2024

Command R+: API Pricing, Context Window & Benchmarks

Command R+ is a language model from Cohere, released in August 2024.

C4AI Command R+ is a 104 billion parameter model with advanced capabilities, including Retrieval Augmented Generation (RAG) and multi-step tool use, optimized for multilingual tasks.

Input
Text
Output
Text

Command R+ benchmarks

Rankings

Quality Tracker

Command R+ Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Fri Aug 07 2026
Notice missing or incorrect data?

Command R+ pricing

Providers

Command R+ starts at $0.250 per million input tokens and $1.00 per million output tokens via Cohere. See all 2 providers below with their per-token pricing, latency, throughput, and modality support.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
Cohere logoCohere
$0.250$1.00128.0K/128.0K
0.65
/
Bedrock logoBedrock
$3.00$15.00128.0K/128.0K
0.50
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Loading chart...
Loading chart...
Loading chart...

Command R+ context window

Input and output token limits for Command R+, plus how it ranks on long-context understanding.

InputOutput
128Ktokens
128Ktokens
192 pages of text
128K
8K128K1M

Command R+ API

Available from the model provider

Command R+ has an official provider API. It is not currently routed through the LLM Stats gateway.

Read the official API documentation

Command R+ latency

Command R+ time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.

Command R+ examples

Recent arena outputs from Command R+, picked from the highest-ranked matchups.

Command R+ license

Command R+ is a proprietary model available under its provider's product and API terms, has 104.0B parameters.

License
CC BY-NC
Hosted access
Parameters
104.0B

Creative Commons Non-Commercial

Command R+ resources

Official sources for Command R+: api documentation, official playground.

Command R+ vs other models

The most-compared alternatives to Command R+ are MiMo-V2.5-Pro, Claude 3 Sonnet, Claude 3 Haiku. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Command R+

Models ranked just above and below Command R+ by LLM Stats score.

 

MiMo-V2.5-Pro

Score pending
 

Claude 3 Sonnet

Score pending
 

Claude 3 Haiku

Score pending
 

Qwen2.5 32B Instruct

Score pending
 

Ministral 3 (8B Base 2512)

Score pending
 

Gemma 2 27B

Score pending

FAQ

Common questions about Command R+.

When was Command R+ released?

Command R+ was released on August 30, 2024 by Cohere. This is the official Command R+ release date tracked on LLM Stats.

How much does Command R+ cost?

Command R+ pricing starts at $0.25 per million input tokens and $1.00 per million output tokens via Cohere, the lowest price among tracked providers.

Is Command R+ available via API?

Yes, Command R+ is available via API. See the official documentation for authentication and endpoint details. It is served by 2 providers tracked on LLM Stats.

How big is Command R+?

Command R+ has 104 billion parameters.

Who created Command R+?

Command R+ was created by Cohere.

What is the license for Command R+?

Command R+ is released under the CC BY-NC license.

What is Command R+ latency?

Command R+ p95 time to first token is 0.50 seconds via Bedrock over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and API workloads.

Where can I use Command R+?

Command R+ is available through 2 providers including Cohere, Bedrock.

What models should I compare Command R+ against?

Common Command R+ comparisons include Command R+ vs MiMo-V2.5-Pro, Command R+ vs Claude 3 Sonnet, Command R+ vs Claude 3 Haiku. Compare them side by side for benchmark scores, pricing, context window, latency and API availability.