- Organizations
- OpenAI
- Whisper Large V3
Whisper Large V3: Benchmarks, Pricing & Examples
Whisper Large V3 is a speech-to-text model from OpenAI, released in November 2024, with a 100M-token context window, and pricing from $0.970 per 1M input tokens.
State-of-the-art multilingual transcription with high accuracy. 10.3% WER, 189x real-time speed. Served by Groq.
Whisper Large V3 pricing
Providers
Whisper Large V3 starts at $0.970 per million input tokens via Groq.
| Provider | Input $/M | Cached input $/M | Output $/M | Context in / out | TTFT p95 s | Output p5 c/s | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.970 | — | — | 100.0M/— | — | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.
Whisper Large V3 context window
Input and output token limits for Whisper Large V3, plus how it ranks on long-context understanding.
Try now
Give your ideas
a voice.
Bring your next story, script, or creative project into your Huggle workspace.
Opens in Huggle
Whisper Large V3 latency
Whisper Large V3 time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.
Whisper Large V3 license
Whisper Large V3 is released under the Apache 2.0 license, which permits commercial use.
- License
- Apache 2.0
- Commercial use allowed
Apache License 2.0 - allows commercial use
FAQ
Common questions about Whisper Large V3.