- Organizations
- MiniMax
- Speech 02 HD
Speech 02 HD: Pricing, Performance & Examples
Speech 02 HD is a text-to-speech model from MiniMax, released in December 2024, with a 3K-token context window, and pricing from $375000000 per 1M input tokens.
High quality production model with emotion support
Speech 02 HD pricing
Providers
Speech 02 HD starts at $375000000 per million input tokens via MiniMax.
| Provider | Input $/M | Output $/M | Context in / out | TTFT p50 / p95 s | Output avg / p5 c/s | Success 7d | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $375000000 | — | 2.5K/— | —/— | —/— | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests. Success is calculated from completed versus failed requests over the trailing seven days.
Speech 02 HD context window
Input and output token limits for Speech 02 HD, plus how it ranks on long-context understanding.
Speech 02 HD API
Run a request to see the response
Use it in your code
OpenAI-compatible endpoint through the LLM Stats gateway.
import requests
response = requests.post(
"https://gateway.llm-stats.com/v1/tts/synthesize",
headers={"Authorization": "Bearer YOUR_API_KEY"},
json={
"model_id": "speech-02-hd",
"text": "Hello, this is a test.",
"format": "mp3",
"sample_rate": 24000,
},
)
with open("output.mp3", "wb") as f:
f.write(response.content)Need an API key? Create one above in the playground, or read the API documentation.
Speech 02 HD latency
Speech 02 HD time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.
Speech 02 HD license
Speech 02 HD is a proprietary model available under its provider's product and API terms.
- License
- Proprietary
- Hosted access
Proprietary license - usage restrictions apply
FAQ
Common questions about Speech 02 HD.