- Organizations
- inclusionai
- Ling 3.1 Flash
Ling 3.1 Flash: Benchmarks, Pricing & Context Window
Ling 3.1 Flash is a language model from InclusionAI, released in September 2026.
Ling 3.1 Flash (Ling-3.1-flash) is InclusionAI / Ant Group's hybrid-reasoning MoE flash model: ~560B total parameters with ~25B active per token. API-only at launch (weights not yet on Hugging Face). Hosted context is commonly 262K
Ling 3.1 Flash benchmarks
Capability tiers
Standing within each category, adjusted for leaderboard depth.
Real tasks performance
High-confidence performance for Ling 3.1 Flash across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.
Performance by conversation depth
How Ling 3.1 Flash holds up as conversations get longer.
Quality Tracker
Ling 3.1 Flash Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
Try now
Make it with
Ling 3.1 Flash.
Ling 3.1 Flash latency
Ling 3.1 Flash time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.
Ling 3.1 Flash examples
Recent arena outputs from Ling 3.1 Flash, picked from the highest-ranked matchups.
Ling 3.1 Flash license
Ling 3.1 Flash has 560.0B parameters.
- Parameters
- 560.0B
Ling 3.1 Flash resources
Official sources for Ling 3.1 Flash: provider documentation, official launch post.
Ling 3.1 Flash vs other models
The most-compared alternatives to Ling 3.1 Flash are Muse Spark 1.3, Qwen3.5-397B-A17B, Atria Dawn Preview. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like Ling 3.1 Flash
Models ranked just above and below Ling 3.1 Flash by LLM Stats score.
FAQ
Common questions about Ling 3.1 Flash.