All tasks
Task·video

Text-to-Video AI Models & Workflows

Describe a shot and get motion. Choose a text-to-video model, drop it into a NodeTool workflow, and generate clips from a prompt alone — on one canvas with your own keys.

B-Roll Reel from a BriefClip on KieExplainer Clip from a ParagraphScript to Narrated Clip

Models for text to video

Veo 3.1
video
Google

Cinematic text-to-video with native synchronized audio.

Sora 2
video
OpenAI

Prompt-to-video with strong physical realism and sound.

Kling 2.6
video
Kling

High-motion generation with strong subject consistency.

Wan 2.6
video
Alibaba

Open-weight text-to-video with fine motion control.

A workflow that does it

Workflow Editor
Note
String Input
A coastal roastery at first light, steam rising off the drum
String Input
handheld documentary, warm morning light, 35mm, shallow depth of field
Prompt
Subject
Look
Wide establishing shot. Subject: {{ subject }} Look: {{ look }} Hold the full scene in frame with a slow push-in. Natural motion only. No t…
Prompt
Subject
Look
Close detail shot of the same scene. Subject: {{ subject }} Look: {{ look }} Tight on texture and material — hands, surfaces, steam. Shallo…
Text To Video
Prompt
fal-ai/ltx-2.3/text-to-video/fast
Text To Video
Prompt
fal-ai/ltx-2.3/text-to-video/fast
Transition
Video A
Video B
Output

Frequently asked questions

What is text-to-video generation?

A model turns a written prompt into a short video clip, inferring subject, motion, and camera work from the text alone — no starting image required.

Which text-to-video model is best?

It depends on the look you want: Veo and Sora lead on cinematic realism and native audio, while open-weight models like Wan let you run locally and tune control.

How do I run it in NodeTool?

Open one of the text-to-video templates below in Studio, connect the provider key, write a prompt, and run the graph.

Do text to video on your own machine

Free, open source, and yours to run. Download Studio, pick a model, and wire it into a workflow with your own keys.