MiniMax · Audio
Speech 2.8
MiniMax's studio-grade speech model with 332 voices, emotion control, voice cloning, and a 50,000-character input limit.
From MiniMax, Speech 2.8 is a cloud text-to-speech model. It converts written text into natural, spoken audio. It supports multiple languages. Voices can be tuned for speed and emotion, and output is available as MP3, WAV, FLAC, and PCM. It runs through Replicate and Runware using your own API key, from $0.10 per 1,000 characters.
Modality
Audio
Available on
ReplicateRunware
Model ID
minimax/speech-2.8
Specs
- Pricing
- $0.10 / 1k chars
- Type
- Text-to-speech
- Languages
- Multilingual
- Voice controls
- Speed, Emotion
- Output formats
- MP3, WAV, FLAC, PCM
About the creator
MiniMax
MiniMax is a Chinese AI lab whose speech and music models are widely used for multilingual voiceover and generative music.
www.minimax.io ↗More from MiniMax