Turn a written scene into an MP4 — multiple models in one workspace, credit cost before you click, refunds if the run fails. Built for marketers and creators who need motion drafts before a shoot exists.
Prompt only — no upload required
Use Text when you have a brief but no plate yet. Skip it when a still already defines the product — that belongs on Image. Reference is for uploaded video or audio guidance, not for empty-handed ideation. The four patterns below are the jobs teams actually run on ToVideo before they open an NLE.
Spin three to five prompt variants of the same offer — change only camera or setting — before you spend on a shoot. Keep duration short and watch the first two seconds muted: if the motion does not read, rewrite the verb, not the adjective pile. Log which variant won so the live shoot brief stays concrete.
Need a hero loop while packaging is unfinished? Text to video gets a watchable placeholder online so design review can move. The moment real key art lands, regenerate with image-to-video from that still so lighting and product shape lock. Until then, treat text outputs as temporary creative, not final brand film.
One prompt per beat: establish → action → reaction. Export each MP4 into an editor timeline and cut on intention. Cramming three scenes into one prompt usually yields mushy mid-shots. If a beat fails, regenerate that beat only — do not restart the whole sequence from a bloated master prompt.
Same prompt, switch 16:9 then 9:16 when the model allows. Regenerating for the target ratio beats forcing a landscape master into a vertical crop that chops faces and logos. Note which models expose which ratios in the dropdown; unavailable options simply do not appear.
Copy a recipe, swap the bracketed parts, generate. One subject, one action, one camera move — then stop. The last panel lists the mistakes that waste the most credits on text to video.
Differentiation is in the workflow: compare models without leaving the form, see cost before you commit, and continue into stills or image-to-video on the same balance.
HappyHorse, Seedance, Veo, Kling, and other text-capable models share one dropdown. Duration, quality, and audio fields refresh to what that model accepts, so invalid combinations never sit in the UI.
Cost updates with model, length, quality, and audio. Failed tasks refund automatically — iterate prompts without silent balance burn. Buy packs or plans on pricing when you need volume.
Same login and credits for AI Image and Video. Ideate in Text, lock a frame in Image when composition matters, then animate with image-to-video so product shape and faces stay consistent.
Close the tab; generation continues on the server. Return through the generator preview or AI Tasks to download the MP4 when status succeeds.
Text = prompt only. Image = animate stills (first frame, or first then last when supported). Reference = image, video, or audio guidance. Each model only enables modes it supports.
Successful runs expose download in the preview panel. Archive outside ToVideo for client delivery, ad platforms, or your editor project.
Product answers for operators shipping drafts — not a glossary of AI buzzwords.
Write one clear scene, check credits, and generate — then change a single variable if you retry.