OpenAI GPT Audio Mini
The cost-efficient tier of OpenAI's speech model, reading text aloud in the same 13 voices at a fraction of the price.
OpenAI's GPT Audio Mini is a cloud text-to-speech model. It converts written text into natural, spoken audio. It offers a choice of 13 voices. It runs through OpenRouter using your own API key, from $0.006 per 1,000 characters.
- Released
- 1 year ago (Aug 2025)
- Pricing
- $0.006 / 1k chars
- Type
- Text-to-speech
- Voices
- 13 to choose from
Examples
Generated with GPT Audio Mini via Runware — the same three scripts read by every text-to-speech model in the catalog, so the only thing that changes between two models’ clips is the model. Each model uses its own default voice; there is no shared voice to hold constant.
Baseline naturalness and pacing
The last train had already left, but she decided to walk anyway. The city was quieter than she remembered, and for the first time in weeks, she wasn't in a hurry.
Emotion, emphasis, and pauses
Wait — you're telling me it actually worked? After all that? I can't believe it. Honestly, I thought we'd lost the whole thing.
Decimals, percentages, version numbers
The API returned 3,481 results in 0.42 seconds, a 12 percent improvement over version 2.5. Latency at the 99th percentile dropped from 840 milliseconds to 610.
OpenAI
OpenAI builds GPT, DALL·E, the Sora family, and the open-weight gpt-oss models, and has been a central force behind the modern wave of generative AI.
openai.com ↗