Text to video AI generator

Describe a shot, get footage back. With sound, in the ratio you need.

Every video model

One prompt box

Written as a prompt, returned as footage

From a sentence to a shot in three steps

  1. 01

    Open Morphic

    Sign up and start creating on a free-flowing, infinite visual canvas.

  2. 02

    Write the shot

    Name the subject, the action, the camera move, and the look. Pick your duration and aspect ratio.

  3. 03

    Generate and compare

    Send the same prompt to a second model, watch both takes, and keep the one that moves better.

Anatomy of a video prompt

A still only needs a subject and a look. A shot also needs to know what moves and how the camera behaves. Name all four and the first take lands much closer.
SubjectActionCameraLook
SubjectA vintage sedan on a coastal roadActionclimbing steadily as the road bendsCameraslow tracking shot from the leftLookwarm low sun, 35mm, soft grain
SubjectA pinball machine mid-gameActionthe ball ricocheting between bumpersCameraclose overhead, locked offLooklit by the cabinet lamps, high shutter speed
SubjectA paper boat in a gutter streamActionspinning once, then carried awayCameralow tracking shot at water levelLookovercast daylight, hand-drawn animation

Use cases

Shots a shoot could not justify

A crane over farmland, a storm rolling in, an aerial you have no permit for. Write the shot instead of chartering it, and see it inside a few minutes.

Ads and social cuts in every ratio

Generate the same idea vertically for a feed and wide for a landing page. Setting the ratio before you generate beats cropping a finished shot afterwards.

Previs and pitch reels

Show the sequence rather than describing it. A handful of generated shots communicates blocking, pace, and tone far faster than a slide of bullet points.

B-roll and connective tissue

Fill the gaps an edit is missing: a hand turning a page, traffic at dusk, steam off a cup. Written to match the cut instead of hunting a library for something close.

All on Morphic

Everything a written shot can turn into

Simple pricing

Get started for free today, with the option to upgrade or cancel anytime.

Basic

$9/ month
billed as $0 per year

1100 monthly credits

1 user only

All models

Workflows

Standard

$24/ month
billed as $0 per year

3625 monthly credits

1 user only

All models

Workflows

Pro

$45/ month
billed as $0 per year

6350 shared monthly credits

1 user

+ up to 4 more at extra cost

All models

Workflows

Pro Max

$170/ month
billed as $0 per year

24650 shared monthly credits

1 user

+ up to 9 more at extra cost

All models

Workflows

Enterprise

For higher limits

Custom

pricing and billing terms

High-volume credits
Custom seat limits
All models
Workflows
Pricing Gradient

Free

For playing around

$0

forever free

Up to 20 credits
1 user only
Limited models
Workflows

FAQs

How does text to video AI work?
You write a shot the way you would brief a camera operator, and the model generates the frames of that shot in sequence. It is not cutting together stock clips, so nothing you get back exists anywhere else. Because the sentence is the entire brief, naming the camera move and the light changes the footage as much as naming the subject does.
Is text to video free to use?
Yes. The free tier lets you write a prompt and generate a clip without paying first, which is enough to see how a model handles your kind of shot. Paid plans lift the volume, unlock the longer and higher-resolution options, and cover commercial use.
Do I need an account to generate a video?
You can watch every example on this page without signing in. Generating your own clip needs a free account, which takes about a minute and no card. It also gives you a place to keep your takes so you can come back and build on them.
How long can a generated clip be?
Individual generations are short by design, in the range of a few seconds, because that is the unit these models work in. Longer pieces get built the way edits always have been: generate the shots you need, then assemble them in order. Thinking in shots rather than in finished films is what makes the results usable.
Can the video have sound?
Yes. Several of the models generate synchronised audio along with the picture, so ambience and effects arrive with the footage rather than being laid on afterwards. If a clip comes back silent or you want something specific, you can generate sound effects, music, or a voiceover separately and place them against the cut.
What should a video prompt include?
Four things carry most of the result: the subject, what it is doing, how the camera behaves, and the look. "A car driving" leaves everything to chance. "A vintage sedan climbing a coastal road at dusk, slow tracking shot from the left, warm low sun, 35mm" gives the model an action, a camera move, a light source, and a finish.
Which model should I use?
They differ in real ways. Some hold physical motion and human movement more convincingly, some are stronger on stylised and animated looks, some return a take faster. Kling 3.0, Hailuo 2.3, Seedance 2.0, Veo 3.1, Vidu Q3, and LTX 2.3 all run in the same workspace, so the fastest answer is to send one prompt to two models and compare the takes.

You might also like