- Organizations
- OpenAI
- Whisper Large V3 Turbo
Whisper Large V3 Turbo: Benchmarks, Pricing & Examples
Whisper Large V3 Turbo is a speech-to-text model from OpenAI, released in November 2024, with a 100M-token context window, and pricing from $0.350 per 1M input tokens.
Fine-tuned version of pruned Whisper Large V3 designed for fast, multilingual transcription. 12% WER, 216x real-time speed. Served by Groq.
Whisper Large V3 Turbo pricing
Providers
Whisper Large V3 Turbo starts at $0.350 per million input tokens via Groq.
| Provider | Input $/M | Cached input $/M | Output $/M | Context in / out | TTFT p95 s | Output p5 c/s | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.350 | — | — | 100.0M/— | — | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.
Whisper Large V3 Turbo context window
Input and output token limits for Whisper Large V3 Turbo, plus how it ranks on long-context understanding.
Try now
Give your ideas
a voice.
Bring your next story, script, or creative project into your Huggle workspace.
Opens in Huggle
Whisper Large V3 Turbo latency
Whisper Large V3 Turbo time to first token, sustained output throughput, and failed-request rate from live model usage over the trailing 7 days.
Whisper Large V3 Turbo license
Whisper Large V3 Turbo is released under the MIT license, which permits commercial use.
- License
- MIT
- Commercial use allowed
MIT License - allows commercial use
FAQ
Common questions about Whisper Large V3 Turbo.