Skip to content
CSuite
Audio · Released Aug 2025

OpenAI GPT Audio

OpenAI's generally available speech model: a chat model that reads text aloud in 13 natural-sounding voices, with consistent voice identity across long passages.

From OpenAI, GPT Audio is a cloud text-to-speech model. It converts written text into natural, spoken audio. It offers a choice of 13 voices. It runs through OpenRouter using your own API key, from $0.12 per 1,000 characters.

Modality
Audio
Available on
Model ID
openai/gpt-audio
Specs
Released
1 year ago (Aug 2025)
Pricing
$0.12 / 1k chars
Type
Text-to-speech
Voices
13 to choose from
Samples

Examples

Generated with GPT Audio via Runware the same three scripts read by every text-to-speech model in the catalog, so the only thing that changes between two models’ clips is the model. Each model uses its own default voice; there is no shared voice to hold constant.

Narration

Baseline naturalness and pacing

The last train had already left, but she decided to walk anyway. The city was quieter than she remembered, and for the first time in weeks, she wasn't in a hurry.

0:00 / 0:10
Expressive

Emotion, emphasis, and pauses

Wait — you're telling me it actually worked? After all that? I can't believe it. Honestly, I thought we'd lost the whole thing.

0:00 / 0:08
Numbers and jargon

Decimals, percentages, version numbers

The API returned 3,481 results in 0.42 seconds, a 12 percent improvement over version 2.5. Latency at the 99th percentile dropped from 840 milliseconds to 610.

0:00 / 0:14
About the creator

OpenAI

OpenAI builds GPT, DALL·E, the Sora family, and the open-weight gpt-oss models, and has been a central force behind the modern wave of generative AI.

openai.com
More from OpenAI

One-time payment. Yours forever.

No subscriptions. No seats. No renewals. Buy CSuite once, future updates included.

Secure checkout via Stripe. Already have a license? Download the app