The AI arena is free today

Open Superagent
StepFunReleased on Feb 2, 2026

Step-3.5-Flash: Benchmarks, Pricing & Context Window

Step-3.5-Flash is a language model from StepFun, released in February 2026, with pricing from $0.100/M input, $0.020/M cached input, $0.400/M output.

Step-3.5-Flash is StepFun's fast, cost-effective text model optimized for quick inference. Built on their Step3 architecture, it offers strong performance across text tasks with low latency and efficient token usage, ideal for production

Input
TextImage
Output
Text

Step-3.5-Flash benchmarks

Capability tiers

Standing within each category, adjusted for leaderboard depth.

Real tasks performance

High-confidence performance for Step-3.5-Flash across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.

Performance by conversation depth

How Step-3.5-Flash holds up as conversations get longer.

Quality Tracker

Step-3.5-Flash Performance Across Datasets

Scores sourced from the model's scorecard, paper, or official blog posts

LLM Stats Logollm-stats.com - Fri Oct 02 2026
Notice missing or incorrect data?

Step-3.5-Flash pricing

Providers

Step-3.5-Flash starts at $0.100 per million input tokens and $0.400 per million output tokens via StepFun. Reused prompt prefixes cost $0.0200 per million cached input tokens.

ProviderInput $/MCached input $/MOutput $/MContext in / outTTFT p95 sOutput p5 c/sModalities in / out
StepFun logoStepFun
$0.100$0.0200$0.40065.5K/8.2K
0.30
—
/

Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.

Step-3.5-Flash model size

Step-3.5-Flash has 196 billion parameters. See how it compares to other models in the same parameter range.

Parameters
196BMoE
Very large (80–200B)
196B
1B7B70B405B

Step-3.5-Flash context window

Input and output token limits for Step-3.5-Flash, plus how it ranks on long-context understanding.

InputOutput
66Ktokens
8Ktokens
≈ 99 pages of text
66K
8K128K1M

Try now

huggle
Step-3.5-Flashin Huggle

Make it with
Step-3.5-Flash.

Step-3.5-Flash

Step-3.5-Flash latency

Step-3.5-Flash time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.

Step-3.5-Flash examples

Recent arena outputs from Step-3.5-Flash, picked from the highest-ranked matchups.

Step-3.5-Flash license

Step-3.5-Flash is released under the Apache 2.0 license, which permits commercial use, has 196.0B parameters.

License
Apache 2.0
Commercial use allowed
Parameters
196.0B

Apache License 2.0 - allows commercial use

Step-3.5-Flash resources

Official sources for Step-3.5-Flash: provider documentation, official playground, official launch post, source repository.

Step-3.5-Flash vs other models

The most-compared alternatives to Step-3.5-Flash are GPT-5.1 Medium, Seed 2.0 Pro, GPT-5.2. Open any pair side-by-side for benchmarks, pricing, context, and latency.

Models like Step-3.5-Flash

Models ranked just above and below Step-3.5-Flash by LLM Stats score.

 

GPT-5.1 Medium

Score pending
 

Seed 2.0 Pro

Score pending
 

GPT-5.2

Score pending
 

GPT-5 Codex

Score pending
 

Qwen3.5-397B-A17B

Score pending
 

Sarvam-105B

Score pending

FAQ

Common questions about Step-3.5-Flash.

When was Step-3.5-Flash released?

Step-3.5-Flash was released on February 2, 2026 by StepFun. This is the official Step-3.5-Flash release date tracked on LLM Stats.

How much does Step-3.5-Flash cost?

Step-3.5-Flash pricing starts at $0.10 per million input tokens, $0.02 per million cached input tokens, $0.40 per million output tokens via StepFun, the lowest price among tracked providers.

How big is Step-3.5-Flash?

Step-3.5-Flash has 196 billion parameters. It ships as an open-weight model, so you can download and run it on your own hardware.

Who created Step-3.5-Flash?

Step-3.5-Flash was created by StepFun.

What is the license for Step-3.5-Flash?

Step-3.5-Flash is released under the Apache 2.0 license. This is an open-source / open-weight license that permits self-hosting.

What is Step-3.5-Flash latency?

Step-3.5-Flash p95 time to first token is 0.30 seconds via StepFun over the trailing 7 days. Lower time to first token means the model begins responding sooner for chat, agents and model workloads.

Where can I use Step-3.5-Flash?

Step-3.5-Flash is available through 1 provider including StepFun.

Where is the Step-3.5-Flash paper or technical report?

Step-3.5-Flash has a paper or technical report available at https://stepfun-ai.github.io/Step-3.5-Flash. Use that source for architecture, training, release and evaluation details.

What models should I compare Step-3.5-Flash against?

Common Step-3.5-Flash comparisons include Step-3.5-Flash vs GPT-5.1 Medium, Step-3.5-Flash vs Seed 2.0 Pro, Step-3.5-Flash vs GPT-5.2. Compare them side by side for benchmark scores, pricing, context window, latency and provider availability.