Narrate a script with a natural voice, score it with generated music, and drop in the sound effects — three kinds of audio generation in one workspace, with a built-in waveform editor for the cleanup. Generation runs on your own API keys; editing runs locally on your machine.
Switch the composer between speech, music, and sound effects — the model picker filters to what supports the active type, and your choice is remembered per category. ElevenLabs for a warm narrator, Lyria for a score, SFX 1.5 for foley: one API key per platform unlocks them all.
Paste the script, pick a model and a voice, and generate. Voices come from each model's own roster, with per-voice controls — speed, stability, similarity — on models that expose them.The clip below the mockup is a real, unretouched Eleven v3 generation from CSuite — press play. The prompt in the bar is the exact script it read.
voiceover.mp3Eleven v30:000:11ReplicateEleven v3Voice · RachelMP3The last train had already left, but she decided to walk anyway. The city was quieter than she remembered…Speak
Real output · Eleven v3narration.mp3
0:00 / 0:10
Compose
Describe the mood, get the music.
Instrumentation, tempo, texture — describe it like a brief to a session musician and the model composes it. Royalty questions don't follow you around: it's your generation, made with your key.The 30-second clip below is a real Lyria 3 composition from the exact prompt shown — the same brief every music model in the catalog answers on its detail page.
rainy-morning.mp3Lyria 30:000:30ReplicateLyria 3Music30sSlow ambient piece for a rainy morning: warm analog pads, soft tape hiss, and a simple four-note piano motif repeating. No drums.Generating
Real output · Lyria 3ambient.mp3
0:00 / 0:30
Foley
Sound effects on demand.
The door slam, the forest ambience, the projector whir — describe the sound instead of digging through sample libraries, and drop the result straight into a video composition or export it for your editor.Below: a real SFX 1.5 generation of the prompt shown — a heavy wooden door in a stone corridor.
door-slam.mp3SFX 1.50:000:10RunwareSFX 1.5Sound effectA heavy wooden door slamming shut in a stone corridor, with a short natural reverb tail.Generate
Real output · SFX 1.5impact.mp3
0:00 / 0:10
Edit
Clean it up without leaving the app.
A waveform editor built into the player covers the usual cleanup — no export to another tool. Every edit renders locally on your machine and saves in place or as a new copy, in lossless WAV.
Trim to the take
Cut the silence off both ends of a recording, or pull one clean take out of a long session — drag the handles on the waveform and everything outside the selection goes.
podcast-take.mp30:06 → 0:48ResetSave new copy
Fade in, fade out
Give a generated music bed a smooth entrance and exit before it goes under a voiceover — set the fade lengths and the envelope is applied to the waveform.
rainy-morning.wavFade in · 1.5sFade out · 2sSave new copy
Fix the level
A quiet voiceover next to a loud music bed is the most common mix problem there is — boost or cut the clip's volume until the levels sit right.
voiceover.mp3Volume120%
Change the speed
Turn a 42-minute lecture recording into a 28-minute listen at 1.5×, or slow a fast take down — the duration updates with the rate.
lecture-recording.mp30.5×1×1.5×2×42:00 → 28:00
Convert to WAV
Hand your editor a lossless file: any import — FLAC, OGG, AAC, MP3 — converts to WAV on save, so downstream tools get clean audio with no stacked compression.
field-notes.flacfield-notes.flac→field-notes.wavWAVSave new copy
Catalog
Available models & providers.
Every audio model in the catalog — text-to-speech, music, and sound effects — all cloud-hosted through Replicate and Runware with your own keys, each with real sample clips on its detail page.
How speech, music, and sound-effect generation work in CSuite — voices, costs, editing, and formats.
For speech: Eleven v3, Eleven Turbo v2.5, Gemini 3.1 Flash TTS, Speech 02 Turbo, and Grok Text-to-Speech. For music: Lyria 3 and Lyria 3 Pro, Eleven Music, and Music 2.6. For sound effects: SFX 1.5. All run in the cloud through Replicate and Runware with your own API keys — browse per-model details and real sample clips at csuite.so/models.
A Type selector switches the composer between the three categories, and the model picker filters to models that support the active one. Your model choice is remembered per category, so flipping from a voiceover to a music bed and back doesn't lose your setup.
Each text-to-speech model ships its own voice roster — pick a voice in the generation settings, and on models that expose them, tune per-voice controls like speed, stability, and similarity. Several models also support multiple languages.
Generation is cloud-only for now — there are no local audio models in the catalog yet. Editing is fully local, though: trims, fades, volume, speed, and conversion all run on your machine, and your recordings never upload anywhere.
Text-to-speech is typically billed per 1,000 characters, and music or sound effects per generation or per second of output — always by the platform (Replicate or Runware) directly to your account, with no CSuite markup or metering. The app's analytics show spend per provider.
Yes. Drop MP3, WAV, FLAC, OGG, or AAC files into your project folder (or drag them into the sidebar) and they open in the same player: trim a take, fade the ends, adjust volume, change speed, or convert — saving in place or as a new copy alongside the original.
Edits render locally through the browser's audio engine and save as WAV — lossless, so repeated edits never stack compression artifacts. Generated clips keep whatever format the model produced (typically MP3), and you can import any of the supported formats.
In your project folder, as plain files on your own disk — generated clips, imports, and edited copies alike. No proprietary library, no cloud sync.
Launch offer · 50% off
One-time payment. Yours forever.
No subscriptions. No seats. No renewals. Buy CSuite once, future updates included.