All tasks
Task·audio

Text-to-Speech AI Models & Workflows

Give your workflow a voice. Pick a text-to-speech model, drop it into a NodeTool graph, and turn scripts into narration and voiceover — with your own keys.

AI SpokespersonLocalise a Script and Revoice ItNarrate a ScriptNarration with a Music Bed

Models for text to speech

ElevenLabs v3
audio
ElevenLabs

Expressive, multi-lingual voices with cloning.

OpenAI TTS
audio
OpenAI

Clean, reliable narration voices.

Kokoro
audio
Local (MLX)

Open-weight TTS that runs on Apple Silicon.

A workflow that does it

Workflow Editor
Note
Video Input
String Input
Our spring release ships today. Faster renders, sharper output, and a price that did not move.
Text To Speech
Text
Audio
inworld/realtime-tts-1.5-max
Lip Sync
Video
Audio
fal-ai/sync-lipsync/v2/pro
Output

Frequently asked questions

What is text-to-speech?

It converts written text into spoken audio using an AI voice — useful for narration, voiceover, and accessibility.

Can I clone or customize a voice?

With providers like ElevenLabs, yes. Swap the TTS node's model and voice settings to change how it sounds.

Does it run offline?

Local models such as Kokoro run entirely in NodeTool Studio on your own machine — no network needed.

Do text to speech on your own machine

Free, open source, and yours to run. Download Studio, pick a model, and wire it into a workflow with your own keys.