Skip to content
CSuite
Audio · Speech-to-text · Released Mar 2023

OpenAI Whisper

OpenAI's hosted Whisper model, which transcribes speech in dozens of languages and returns timed segments for subtitles.

OpenAI's Whisper is a cloud speech-to-text model. It transcribes spoken audio into written text. It handles multiple spoken languages and can return timestamped segments along with the transcript. It also supports an optional text prompt that guides the spelling of names and terms. It runs through OpenAI using your own API key, from about $0.006 per minute of audio.

Modality
Audio
Available on
Model ID
openai/whisper-1
Specs
Released
3 years ago (Mar 2023)
Pricing
$0.006 per minute of audio
Type
Speech-to-text
Languages
Multilingual
Timestamps
Timed segments
About the creator

OpenAI

OpenAI builds GPT, DALL·E, the Sora family, and the open-weight gpt-oss models, and has been a central force behind the modern wave of generative AI.

openai.com ↗
More from OpenAI

One-time payment. Yours forever.

No subscriptions. No seats. No renewals. Buy CSuite once, future updates included.

Secure checkout via Stripe. Already have a license? Download the app