Skip to main content
Audio nodes produce speech and audio assets for ads, Hyperframes voiceovers, URL-to-video handoffs, and standalone Studio video pipelines. Add an audio node to any canvas, choose a mode, and connect the output to downstream video or export nodes.

What audio modes are available?

Convert a script to natural-sounding speech. Paste or type your text, choose a voice, and generate.Default engine: ElevenLabs V3. Gemini TTS is also available as an alternative.Best for: voiceovers, ad narration, product explainers, Hyperframes scripts.

Which voice providers are supported?

See Audio models for a full list of available voices, languages, and credit costs per mode.

How do you use audio in a video pipeline?

1

Add an audio node

Open the toolbar and add an Audio node. Choose TTS or Dialogue mode.
2

Write your script

Type or paste the script into the text field. For Dialogue mode, label each speaker turn.
3

Select a voice

Use the Voice selector mode to preview options, then switch back to TTS and apply your chosen voice.
4

Generate

Click Run. The audio renders as a waveform output on the node.
5

Connect to video or export

Drag the audio output to a lip sync node, a Hyperframes voiceover slot, or an export/merge node.
For Hyperframes, generate voiceover at the Voiceover stage rather than attaching a separate audio node — Hyperframes handles timing alignment automatically when you use the built-in voiceover step.

What’s next?

Hyperframes

Use voiceover in structured animated video

Video nodes

Lip sync audio to a face video

Audio models

Full voice and model reference