The AI arena is free today

Open Superagent

Nemotron 3 Super (120B A12B) vs Parse

Comparing Nemotron 3 Super (120B A12B) and Parse across benchmarks, pricing, and capabilities.

NVIDIA · Cohere · Updated for 2026

Which is better?

Nemotron 3 Super (120B A12B) and Parse trade strengths across price, capabilities, and technical limits. The better choice depends on the workload.

Nemotron 3 Super (120B A12B) also accepts a larger context window (262,144 input tokens), making it the stronger choice for long documents and large codebases.

Based on current LLM Stats indexes, shared benchmarks, pricing, and model metadata for 2026.

Choose Nemotron 3 Super (120B A12B)

  • you process long inputs — it offers a 262,144 token context window
  • you need open weights you can self-host or fine-tune

Choose Parse

  • you want the most recent training data — it shipped Aug 2026

At a glance

The differences that matter most.

Benchmark wins
Input price
$0.10 / M
— / M
Output price
$0.50 / M
— / M
Context window
262,144
8,192

Individual benchmarks

23 reported for Nemotron 3 Super (120B A12B) · 1 for Parse

No common benchmarks found

Nemotron 3 Super (120B A12B) and Parsedon't have any common benchmark datasets to compare. They may have been evaluated on different testing suites.

Human preference

Blind head-to-head votes and playground preference scores

Model Size

Parameter count comparison

117.7B diff

Nemotron 3 Super (120B A12B) has 117.7B more parameters than Parse, making it 5117.4% larger.

NVIDIA
Nemotron 3 Super (120B A12B)
120.0Bparameters
Cohere
Parse
2.3Bparameters
120.0B
Nemotron 3 Super (120B A12B)
2.3B
Parse

Context Window

Maximum input and output token capacity

Nemotron 3 Super (120B A12B) accepts 262,144 input tokens compared to Parse's 8,192 tokens. Only Nemotron 3 Super (120B A12B) specifies output context (262,144 tokens).

NVIDIA
Nemotron 3 Super (120B A12B)
Input262,144 tokens
Output262,144 tokens
Cohere
Parse
Input8,192 tokens
Output- tokens
Fri Sep 04 2026 • llm-stats.com

Input capabilities

Documented input modalities across available providers

Parse supports multimodal inputs, whereas Nemotron 3 Super (120B A12B) does not.

Parse can handle both text and other forms of data like images, making it suitable for multimodal applications.

Nemotron 3 Super (120B A12B)

Text
Images
Audio
Video

Parse

Text
Images
Audio
Video

License

Usage and distribution terms

Nemotron 3 Super (120B A12B) is licensed under NVIDIA Open Model License Agreement , while Parse uses a proprietary license.

License differences may affect how you can use these models in commercial or open-source projects.

Nemotron 3 Super (120B A12B)

NVIDIA Open Model License Agreement

Open weights

Parse

Proprietary

Closed source

Release Timeline

When each model was launched

Nemotron 3 Super (120B A12B) was released on 2026-03-11, while Parse was released on 2026-08-27.

Parse is 6 months newer than Nemotron 3 Super (120B A12B).

Nemotron 3 Super (120B A12B)

Mar 11, 2026

5 months ago

Parse

Aug 27, 2026

1 weeks ago

5mo newer

Knowledge Cutoff

When training data ends

Nemotron 3 Super (120B A12B) has a documented knowledge cutoff of 2025-06-01, while Parse's cutoff date is not specified.

We can confirm Nemotron 3 Super (120B A12B)'s training data extends to 2025-06-01, but cannot make a direct comparison without Parse's cutoff date.

Nemotron 3 Super (120B A12B)

Jun 2025

Parse

Provider Availability

Nemotron 3 Super (120B A12B) is available from DeepInfra. Parse is available from Azure, Cohere.

Nemotron 3 Super (120B A12B)

deepinfra logo
Deepinfra
Input Price:Input: $0.10/1MOutput Price:Output: $0.50/1M

Parse

azure logo
Azure
cohere logo
Cohere
* Prices shown are per million tokens

Outputs Comparison

Notice missing or incorrect data?Start an Issue discussion

Judge for yourself.

Run your own prompts against Nemotron 3 Super (120B A12B) and Parse side-by-side, then vote on the output you prefer.

Nemotron 3 Super (120B A12B)
✓ Preferred
Parse
Open in Playground

FAQ

Common questions about Nemotron 3 Super (120B A12B) vs Parse.

Which is better, Nemotron 3 Super (120B A12B) or Parse?

Nemotron 3 Super (120B A12B) (NVIDIA) and Parse (Cohere) each have strengths in different areas. Compare their benchmark scores, pricing, context windows, and capabilities above to determine which fits your needs.

How does Nemotron 3 Super (120B A12B) compare to Parse in benchmarks?

Nemotron 3 Super (120B A12B) scores HMMT 2025: 94.7%, RULER: 91.8%, AIME 2025: 90.2%, WMT24++: 86.7%, MMLU-Pro: 83.7%. Parse scores ParseBench: 79.2%.

What are the context window sizes for Nemotron 3 Super (120B A12B) and Parse?

Nemotron 3 Super (120B A12B) supports 262K tokens and Parse supports 8K tokens. A larger context window lets you process longer documents, conversations, or codebases in a single request.

What are the main differences between Nemotron 3 Super (120B A12B) and Parse?

Key differences include context window (262K vs 8K), multimodal support (no vs yes), licensing (NVIDIA Open Model License Agreement vs Proprietary). See the full comparison above for benchmark-by-benchmark results.

Who makes Nemotron 3 Super (120B A12B) and Parse?

Nemotron 3 Super (120B A12B) is developed by NVIDIA and Parse is developed by Cohere.