Skip to content
CSuite
Features · Video

Generate clips. Cut them into a film.

Single clips from a prompt, a photo, or start- and end-frames — with the frontier model you pick per shot. Then trim, crop, and convert clips, or assemble full videos on a multi-track timeline with AI drafting the scenes. Cloud generation runs on your own API keys, open models render locally through the GGML runtime, and editing and rendering run on your machine.

Choose

Pick the right model per shot.

Veo 3.1 for a hero shot with synced audio, Seedance 2.0 Fast for cheap iterations, Kling 3.0 or Gen-4.5 when a scene needs a different look. One API key per platform unlocks every model it hosts, and your clips, prompt history, and settings stay in the same workspace no matter which model renders.
Generate

From prompt to footage.

Describe the shot — subject, camera move, light — pick an aspect ratio and resolution, and render. Long jobs stream progress into the composer, and finished clips land in your project folder like any other file.The clip playing here is a real, unretouched Veo 3.1 generation from CSuite — the exact prompt is in the bar below it, and the same test renders for every video model on its catalog page.
Animate

Start from a still, end in motion.

Attach a photo as the start frame and describe the motion — the model animates from exactly that image, keeping your composition and subject. On models that accept an end frame too, give it the first and last shot and it generates the move that connects them: perfect for product turns and scene transitions.
Enhance

Type the gist, get the direction.

Press the wand and one line becomes a proper shot: subject and action, camera movement, framing, light, and mood. With a start frame attached it directs the motion instead of re-describing the scene, and it only writes sound when the model actually renders audio.Video is billed per second and takes minutes to come back, so the vague attempt is the costly one.
Edit

Fix clips without leaving the app.

The everyday cuts that usually mean opening an editor, built in and powered by ffmpeg on your machine — with progress you can cancel, and nothing uploaded anywhere.

Resize for the destination

Drop a full-res render down to 720p for a web embed or 480p for a quick share — pick a preset height from 240p to 1080p and the width follows the clip's aspect ratio.

Crop to a new aspect ratio

Turn a 16:9 landscape shot into a 9:16 vertical for Shorts and Reels, or a 1:1 square for a feed — crop by aspect ratio and keep the action centered.

Convert the container

Deliver the same clip as MP4, WebM, or MOV without a round-trip through another tool — a WebM for your site, a MOV for the editor who asked for one. Or export a looping GIF at the frame rate and width you want — it lands in your Image workspace, ready to drop into a README or a Slack thread.
TrimSet an in and out point by typing them or parking the playhead and clicking Set. The kept range shows as a band on the scrubber.
SpeedSlow a clip to half speed or push it to 2×, with the pitch of any audio preserved rather than chipmunked.
Audio in or outPull a clip's soundtrack out as its own MP3 — it appears in the Audio workspace — or strip it for a silent cut, which takes seconds since the picture isn't re-encoded.
Grab a frameSave the frame under the playhead as a PNG, or send it straight to the composer as the start frame of the next clip. That's how you extend a shot.
UpscaleTake a clip to 1080p or 4K with Real-ESRGAN — no prompt, just pick it from the AI menu. Runs on your Replicate key.
Remove the backgroundMatte a subject out to a green screen with Robust Video Matting. Built for talking-head and lip-sync footage.
Step frame by frameArrow keys move the playhead a second at a time; hold Shift and it steps a single frame, for finding the exact cut.
Know what you're holdingOne panel with duration, resolution, frame rate, codecs, and bitrate — plus the model and prompt behind anything CSuite generated.
Compositions

Not just clips — whole videos.

A second creation mode for long-form work: assemble titles, images, clips, and audio on a multi-track timeline and render one finished video. Describe it in plain English and the AI drafts the whole thing for you.

A multi-track timeline

Titles, images, video, and audio each get their own track. Drag to move, pull an edge to resize, drag up or down to change what draws on top, and snap cleanly to neighbouring clips, the playhead, or a clip's own true length when you're trimming it. Fades, volume, motion, and exact position live in the layer panel — or just drag and scale a title directly on the preview, with guides that snap it to the centre and the edges.

AI drafts the whole cut

Tell the assistant what the video is for and it plans the beats and lays them onto the timeline as real clips — titles, voice-over, music, sound effects — then generates the image and audio layers in parallel, so timings and title cards are playable immediately while the rest backfill. It picks one visual theme and holds every image to it, so the shots look like one shoot. Refine with follow-ups: nudge a timing, reword a title, quieten the music. Each request touches only what you asked about and undoes in one step.

Turn a still into a shot

Every layer has a properties panel: timing and fades, exact position and size, motion, volume and looping, and for titles the fonts actually installed on your machine. Image layers get one more — Make video. Describe the motion and your chosen model animates the still from that exact frame; the clip swaps in place, keeping its timing and framing, and one ⌘Z puts the image back. Changed the shape of the whole video? Switch aspect ratio and every clip is rescaled to fit rather than stretched.

Render on your machine

Export at up to 4K and 60 fps as MP4, WebM, or MOV. The render runs locally with ffmpeg — your footage never uploads — and the finished file lands in your project folder like any other clip.
Catalog

Available models & providers.

Every video model in the catalog, in one picker. Cloud models run through Runware and Replicate with your own keys; open-weight models run locally through the GGML runtime — each with real sample clips on its detail page.

RunwareCloud · your API key
ReplicateCloud · your API key
GGMLLocal · your hardware
Veo 3.1
Cloud
Google's high-quality video model with native audio, reference-image support, and frame-to-frame control.
Runware · Replicate
Veo 3.1 Fast
Cloud
Google's fast video model for generating high-quality short clips from text prompts.
Runware · Replicate
Veo 3.1 Lite
Cloud
Google's cost-efficient Veo tier, generating short clips with native audio at roughly half the price of Veo 3.1 Fast.
Runware · Replicate
Gen-4.5
Cloud
Runway's flagship video model with realistic physics and strong visual fidelity for text-to-video and image-to-video.
Runware · Replicate
Kling 3.0
Cloud
Kuaishou's cinematic video model generating clips up to 1080p with native audio, lip sync, and sound effects from text or images.
Runware · Replicate
Gemini Omni Flash
Cloud
Google's fast, cost-efficient omni video model for text-to-video, image-to-video, and video-to-video with reference and frame conditioning.
Runware
Gemini Omni Flash 1.1
Cloud
Google's updated omni video model, adding a 360p draft tier and 4K output, start-to-end keyframe interpolation, reference videos, and natively synchronized audio.
Runware · Replicate
Flux 3 Video
Cloud
Black Forest Labs' multimodal video model: text, image, or video in, with synchronized audio, lip-sync, and a cheap draft mode.
Runware · Replicate
LTX-2.5 Pro
Cloud
Lightricks' high-fidelity video model — native synchronized audio, first-to-last-frame control, and directable camera moves at up to 1080p.
Runware
LTX-2.5 Fast
Cloud
Lightricks' fast video model — synchronized audio, first-to-last-frame control, and clips up to 20 seconds, at resolutions up to 4K.
Runware · Replicate
H3
Cloud
MiniMax's Hailuo 3 video model with native synchronized audio, multi-reference consistency, and prompt-driven editing of a finished clip.
Runware · Replicate
H3 Max
Cloud
MiniMax's throughput-tuned H3, built for faster clips with native synchronized audio and first-to-last frame control. It trades H3's multi-reference conditioning and 2K tier for speed.
Runware
Grok Imagine Video 1.5
Cloud
xAI's image-to-video model — animates a still image into a short clip with synchronized audio.
Runware · Replicate
HappyHorse 1.1
Cloud
Alibaba's video model with three modes — text-to-video, image animation, and reference-to-video from up to 9 images.
Runware · Replicate
Seedance 2.5
Cloud
ByteDance's flagship multimodal video model — native synchronized audio, clips up to 30 seconds, and large reference sets for character and scene consistency.
Runware · Replicate
Seedance 2.0
Cloud
ByteDance's advanced video model with improved motion coherence and scene quality.
Runware · Replicate
Seedance 2.0 Mini
Cloud
ByteDance's lightweight Seedance 2.0 tier — about half the cost per second, 480p/720p with synchronized audio.
Runware · Replicate
Seedance 2.0 Fast
Cloud
ByteDance's low-latency Seedance 2.0 tier — quick 480p/720p generation with synchronized audio.
Runware · Replicate
Seedance 1.5 Pro
Cloud
ByteDance's cinematic Seedance 1.5 tier — native audio-visual generation with lip-sync and strong camera control, up to 1080p.
Runware · Replicate
Seedance 1 Lite
Cloud
Lightweight Seedance model for quick, stylised video generation.
Replicate
Wan 2.7
Cloud
Alibaba's open-weight Wan 2.7 video model (27B MoE) generating clips with synchronized audio from text or images.
Runware
Wan 3.0
Cloud
Alibaba's all-in-one Wan 3.0 video model, generating clips of up to 30 seconds with native audio from text, frame images, or a deep bench of reference media.
Runware
Wan 3.0 Prime
Cloud
Alibaba's lower-latency Wan 3.0 tier, with the same multimodal workflows, 30-second ceiling, and native audio as the base model.
Runware
Wan 2.2 TI2V-5B
Local
Alibaba's compact 5B video model, running fully offline on your own machine. Handles both text-to-video and image-to-video, and needs no API key or network once the weights are downloaded.
GGML
Wan 2.1 T2V 1.3B
Local
Alibaba's smallest Wan video model, and the lightest local option. Runs at 480p on modest hardware and shares its text encoder with the other Wan entries, so it costs little extra once one of them is installed.
GGML
Wan 2.2 T2V A14B
Local
Alibaba's largest open Wan model, a two-expert mixture that swaps between a high-noise and a low-noise transformer partway through denoising. Highest quality of the local Wan entries and by far the heaviest.
GGML
LTX-2.3 Distilled
Local
Lightricks' 22B video model, distilled for fewer sampling steps. Generates synchronised audio alongside the video, and drives a Gemma 3 language model as its text encoder.
GGML
H3 Local
Local
MiniMax's H3 video model running offline, with synchronised audio. The heaviest local entry by a wide margin: its text encoder alone is a 32B vision-language model.
GGML
FAQ

Video questions, answered.

How video generation, clip editing, and compositions work in CSuite — models, sound, costs, and rendering.

One-time payment. Yours forever.

No subscriptions. No seats. No renewals. Buy CSuite once, future updates included.

Secure checkout via Stripe. Already have a license? Download the app