One prompt in, one clip out

Free Text to Video AI from one written prompt

Text to video AI turns a written prompt into a short video clip. Describe what is in the shot, what happens, how the camera moves and how it looks, and Seedance 1.5 Pro renders it, usually in about two minutes. Free to start.

About 2 minutes

Typical clip, measured

5 seconds

Per clip from this page

Words only

No photo needed

What is text to video AI?

Text to video AI is a video model that makes a clip from words alone: you describe the shot, and it draws every frame, the movement and the light. No photo, footage or editing app is needed. Here it runs on Seedance 1.5 Pro, the fast Seedance version that works from text.

It runs in the browser, inside the Studio of the free AI video generator. Video needs a free account, and finished clips land in Generations, where you can watch them and share them by link.

How do you write a text to video prompt?

Name four things in plain words: what is in the shot and what it does, where the camera is, how fast things move, and the look. One action fits a 5 second clip best.

1
Subject and action

One subject doing one thing: "a red fox crouches, then dives into deep snow". Three actions in one clip get rushed or dropped.

2
Camera

Say where the camera is and how it moves: crane down, tracking shot, push-in, orbit, FPV, static. A static camera keeps fine detail sharp.

3
Motion and speed

Words like slow motion, drifts, races or wobbles set the pace. Slow motion stretches a split second, a splash or a shatter, over the whole clip.

4
Look and light

End with the look, anime, claymation or documentary, and the light: low winter sun, one hard spotlight, dusk. They change a clip more than adjectives do.

How long does a text to video clip take?

We rendered the eight clips on this page on 9 Oct 2026 on Seedance 1.5 Pro at 5 seconds. Each was ready 1 min 31 s to 6 min 08 s after the send. The prompts ran from 39 to 55 words.

ClipTechniqueWords in the promptReady in
01Crane down to the red umbrellaCrane shot511 min 59 s
02FPV dive down a waterfallFPV drone401 min 38 s
03Rising with the jellyfishRising camera391 min 34 s
04Fox dive in slow motionSlow motion, tracking456 min 08 s
05Glass king shattersSlow motion close-up392 min 41 s
06Ink blooming in waterMacro, static camera473 min 48 s
07Clay octopus on drumsClaymation551 min 31 s
08Lantern river, anime lookAnime style, pan441 min 39 s
Send to finished clip
Crane down to the red umbrella
1 min 59 s
FPV dive down a waterfall
1 min 38 s
Rising with the jellyfish
1 min 34 s
Fox dive in slow motion
6 min 08 s
Glass king shatters
2 min 41 s
Ink blooming in water
3 min 48 s
Clay octopus on drums
1 min 31 s
Lantern river, anime look
1 min 39 s
Measured on 9 Oct 2026: eight 5 second clips on Seedance 1.5 Pro.

How to make a video from text

Three steps, all on this page.

1
Write the shot

Subject, action, camera and look, in one to three sentences. Or start from a technique under the box.

2
Send it

Sign in when asked. Seedance 1.5 Pro renders the clip, usually in about two minutes.

3
Change one thing

The clip lands in Generations. Change the camera or the light in the prompt and send again for a new take.

Questions about text to video AI

What people ask before their first clip.

Yes, it is free to start, with no subscription. Video uses more energy than pictures; energy refills when you share, post or invite friends, and a daily bonus tops it up.

Type a prompt in the box above: what is in the shot, what happens, how the camera moves and the look. Send it, sign in, and the clip arrives in Generations, usually in about two minutes.

One to three sentences. The prompts on this page run from 39 to 55 words. A longer prompt does not make a longer clip: 5 seconds holds one subject and one action.

Push-in, pull-back, pan, tilt, orbit, crane, tracking shot, FPV and a static camera all work in plain words. Name one move per clip.

Seedance 1.5 Pro makes 5, 10 or 12 second clips. This page sends 5 seconds; pick another length in the Studio settings in the chat.

The model fills in whatever the prompt leaves open. Name the camera and the light, keep to one action, and send again: every send is a new take.

Yes, with image to video: the photo becomes the first frame and the prompt only has to describe the motion.

No. It runs in the browser on a phone or a computer. Video needs a free account.

Reviewed by Arthur Down, editor · Updated · How we check facts