5 steps. Every video. Every time.
The guided workflow builder turns chaotic creative requests into a repeatable video factory. Each step enforces platform constraints and tracks completion. Save your progress with Ctrl+S.
Format & Layout
Choose a layout style — Hyper-Realistic (powered by Seedance 2), Product Marketing or Multi-Shot (both powered by Kling 3.0 Omni Pro), or Custom (pick any AI model directly, including Sora 2 Pro for realistic UGC-style footage). Each layout locks in a dedicated AI model and constrains format options to what that model supports best.
Once you select a layout, aspect ratio, resolution, and duration options automatically adjust to the assigned model's capabilities. The Custom layout exposes all 10 AI models for power users who want full control.
Character & Product
Choose from a global library of AI presenters or generate a custom character. Add product placement with compositing or 3D rotating display styles.
Custom character generation uses AI image generation (5 credits). Select products from your brand's product library and configure placement position and animation.
Background & Environment
AI-generated backgrounds with custom prompts or preset environments. Configure color grading, camera angles, and scene composition.
Choose AI-generated or preset environments and set lighting, color treatment, and style references for the overall visual aesthetic. (Brand-kit enforcement — color palette and voice-and-tone — is applied later, at step 4.)
Concept & Script
Write or paste your script, then break it into scenes with speakers and duration hints. Live word count and estimated duration keep you on track.
Each scene can have a dedicated speaker, text cues, and timing hints. The system estimates total video duration based on speech rate and scene count.
Audio
Set voice and background music preferences in a single step — source, language, accent, tone, pacing, and a music description with a narration/music volume balance. These are passed as hints to the selected video model, which may honor them in the generated output.
Per-speaker voice settings with pacing (Slow, Normal, Fast), emphasis words, pause markers, and a pronunciation guide for brand terms, plus a music description (mood and genre) — all provided as prompt guidance to the chosen video model. Whether spoken audio or music appears in the final clip depends on the model's audio capability.