The AI arena is free today

Open Superagent

Command R+ vs Phi 4 Reasoning Plus

Phi 4 Reasoning Plus leads the LLM Stats Score 15.4 to 0.2.

Cohere · Microsoft · Updated for 2026

Which is better?

Phi 4 Reasoning Plus leads the overall LLM Stats Score 15.4 to 0.2, ranking #209 overall.

Based on current LLM Stats indexes, shared benchmarks, pricing, and model metadata for 2026.

Choose Command R+

  • you want predictable pricing at $0.25/M input and $1.00/M output

Choose Phi 4 Reasoning Plus

  • overall performance matters — it scores 15.4 and ranks #209 on LLM Stats
  • your work emphasizes reasoning — it leads those capability indexes
  • you want the most recent training data — it shipped Apr 2025

At a glance

The differences that matter most.

Core performance indexes
0.2
#305
15.4
#209
-0.1
#301
15.6
#204
Cost, coverage & limits
Benchmark wins
Input price
$0.25 / M
— / M
Output price
$1.00 / M
— / M
Context window
128,000

Capability indexes

Additional strengths measured across groups of related public benchmarks

1 shared
Index
Command R+
Phi 4 Reasoning Plus
-1.7#292
18.3#167
Conservative TrueSkill rating · higher is betterHow scores work

Individual benchmarks

6 reported for Command R+ · 11 for Phi 4 Reasoning Plus

No common benchmarks found

Command R+ and Phi 4 Reasoning Plusdon't have any common benchmark datasets to compare. They may have been evaluated on different testing suites.

Human preference

Blind head-to-head votes and playground preference scores

Model Size

Parameter count comparison

90.0B diff

Command R+ has 90.0B more parameters than Phi 4 Reasoning Plus, making it 642.9% larger.

Cohere
Command R+
104.0Bparameters
Microsoft
Phi 4 Reasoning Plus
14.0Bparameters
104.0B
Command R+
14.0B
Phi 4 Reasoning Plus

Context Window

Maximum input and output token capacity

Only Command R+ specifies input context (128,000 tokens). Only Command R+ specifies output context (128,000 tokens).

Cohere
Command R+
Input128,000 tokens
Output128,000 tokens
Microsoft
Phi 4 Reasoning Plus
Input- tokens
Output- tokens
Sat Aug 29 2026 • llm-stats.com

License

Usage and distribution terms

Command R+ is licensed under CC BY-NC, while Phi 4 Reasoning Plus uses MIT.

License differences may affect how you can use these models in commercial or open-source projects.

Command R+

CC BY-NC

Open weights

Phi 4 Reasoning Plus

MIT

Open weights

Release Timeline

When each model was launched

Command R+ was released on 2024-08-30, while Phi 4 Reasoning Plus was released on 2025-04-30.

Phi 4 Reasoning Plus is 8 months newer than Command R+.

Command R+

Aug 30, 2024

2.0 years ago

Phi 4 Reasoning Plus

Apr 30, 2025

1.3 years ago

8mo newer

Knowledge Cutoff

When training data ends

Phi 4 Reasoning Plus has a documented knowledge cutoff of 2025-03-01, while Command R+'s cutoff date is not specified.

We can confirm Phi 4 Reasoning Plus's training data extends to 2025-03-01, but cannot make a direct comparison without Command R+'s cutoff date.

Command R+

Phi 4 Reasoning Plus

Mar 2025

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion

Judge for yourself.

Run your own prompts against Command R+ and Phi 4 Reasoning Plus side-by-side, then vote on the output you prefer.

Command R+
✓ Preferred
Phi 4 Reasoning Plus
Open in Playground

FAQ

Common questions about Command R+ vs Phi 4 Reasoning Plus.

Which is better, Command R+ or Phi 4 Reasoning Plus?

Phi 4 Reasoning Plus leads the LLM Stats Score 15.4 to 0.2. Command R+ is made by Cohere and Phi 4 Reasoning Plus is made by Microsoft. The best choice depends on your use case — compare their capability indexes, individual benchmarks, pricing, and limits above.

How does Command R+ compare to Phi 4 Reasoning Plus in benchmarks?

Command R+ scores HellaSwag: 88.6%, Winogrande: 85.4%, MMLU: 75.7%, ARC-C: 71.0%, GSM8k: 70.7%. Phi 4 Reasoning Plus scores FlenQA: 97.9%, HumanEval+: 92.3%, IFEval: 84.9%, OmniMath: 81.9%, AIME 2024: 81.3%.

What are the context window sizes for Command R+ and Phi 4 Reasoning Plus?

Command R+ supports 128K tokens and Phi 4 Reasoning Plus supports an unknown number of tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between Command R+ and Phi 4 Reasoning Plus?

Key differences include LLM Stats Score (0.2 vs 15.4), licensing (CC BY-NC vs MIT). See the full comparison above for benchmark-by-benchmark results.

Who makes Command R+ and Phi 4 Reasoning Plus?

Command R+ is developed by Cohere and Phi 4 Reasoning Plus is developed by Microsoft.