- Organizations
- inclusionai
- Ling 3.0 Flash Fin
Ling 3.0 Flash Fin: API Pricing, Context Window & Benchmarks
Ling 3.0 Flash Fin is a language model from Unknown Organization, released in September 2026, with a 262K-token context window, and pricing from $0.060/M input, $0.012/M cached input, $0.180/M output.
Ling-3.0-flash-Fin is the first finance-enhanced model in the Ant Ling family. Developed by Ant Group with leading financial institutions and domain experts, it extends Ling-3.0-flash through continued training on high-quality financial
Ling 3.0 Flash Fin benchmarks
Capability tiers
Standing within each category, adjusted for leaderboard depth.
Real tasks performance
High-confidence performance for Ling 3.0 Flash Fin across real-world prompt categories. Only 95% intervals at most 4 points wide are shown.
Performance by conversation depth
How Ling 3.0 Flash Fin holds up as conversations get longer.
Quality Tracker
Ling 3.0 Flash Fin Performance Across Datasets
Scores sourced from the model's scorecard, paper, or official blog posts
Ling 3.0 Flash Fin pricing
Providers
Ling 3.0 Flash Fin starts at $0.0600 per million input tokens and $0.180 per million output tokens via DeepInfra. Reused prompt prefixes cost $0.0120 per million cached input tokens.
| Provider | Input $/M | Cached input $/M | Output $/M | Context in / out | TTFT p95 s | Output p5 c/s | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.0600 | $0.0120 | $0.180 | 262.1K/262.1K | — | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.
Ling 3.0 Flash Fin context window
Input and output token limits for Ling 3.0 Flash Fin, plus how it ranks on long-context understanding.
Ling 3.0 Flash Fin API
Available from the model provider
Ling 3.0 Flash Fin is available from DeepInfra. It is not currently routed through the LLM Stats gateway.
Read the official API documentationLing 3.0 Flash Fin latency
Ling 3.0 Flash Fin time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.
Ling 3.0 Flash Fin examples
Recent arena outputs from Ling 3.0 Flash Fin, picked from the highest-ranked matchups.
Ling 3.0 Flash Fin license
Ling 3.0 Flash Fin has 124.0B parameters.
- Parameters
- 124.0B
Ling 3.0 Flash Fin resources
Official sources for Ling 3.0 Flash Fin: api documentation, model weights.
Ling 3.0 Flash Fin vs other models
The most-compared alternatives to Ling 3.0 Flash Fin are Qwen3.7 Max, Kimi K2.6, Muse Spark 1.1. Open any pair side-by-side for benchmarks, pricing, context, and latency.
Models like Ling 3.0 Flash Fin
Models ranked just above and below Ling 3.0 Flash Fin by LLM Stats score.
FAQ
Common questions about Ling 3.0 Flash Fin.