All tasks
Task·audio

Text-to-Music AI Models & Workflows

Describe a mood and get a track. Choose a text-to-music model, drop it into a NodeTool workflow, and score videos or generate loops — with your own keys.

Music Video VisualizerMeeting Transcript SummarizerSummarize AudioTranscribe Audio

Models for text to music

Suno v4
audio
Suno

Full songs with vocals from a text prompt.

Udio
audio
Udio

High-fidelity music generation and extension.

Stable Audio
audio
Stability

Prompt-driven instrumental tracks and loops.

A workflow that does it

Workflow Editor
Note
Audio Input
String Input
electronic ambient
String Input
abstract geometric patterns, neon colors, flowing energy
String Input
8
Automatic Speech Recognition
Audio
Text
openai/whisper-large-v3
Prompt
Transcription
Genre
Style
Agent
Prompt
Text
gpt-5-mini
Prompt
Analysis
Count
Genre
Style
List Generator
Prompt
gpt-5-mini
For Each
Input List
Text To Image
Prompt
fal-ai/flux/schnell
Collect
Input Item
Frame To Video
Frame
Add Audio
Video
Audio
Output

Frequently asked questions

What is text-to-music generation?

A model composes original audio from a written prompt — you describe genre, mood, and instruments, and it returns a track.

Can I score a video with it?

Yes. Generate a track from the prompt, then mix it under your footage with an add-audio node in the same graph.

Are the tracks royalty-free?

That depends on the provider's license — check the model's terms. NodeTool just wires the model in; you bring your own key.

Do text to music on your own machine

Free, open source, and yours to run. Download Studio, pick a model, and wire it into a workflow with your own keys.