Nodaro Docs
DocumentationNode ReferenceModelsAI Agents (MCP)DevelopersSelf-hostingResearch
Audio

Audio nodes

Every audio node in Nodaro, from text to speech, voice changing and dubbing to Suno music, sound effects, stem separation, editing, sync and transcription.

The Audio tab of the node picker holds every node that makes, changes, cleans, edits or reads sound. It covers speech and voiceover, voice changing and dubbing, music, sound effects, stem separation, audio editing, recording sync and transcription. Press Tab on the canvas and open the Audio tab to see the nodes, grouped in nine families.

The nine families

Add Your Own

Bring your own sound into a workflow. Upload Audio takes a file or a direct link and lets you trim it before upload. Reference Audio also extracts the sound track of a YouTube video.

Speech & Voiceover

Turn text into speech with ElevenLabs voices. Text to Speech reads a text with one voice, in up to 46 languages, with audio tags for emotion. Text to Dialogue voices a whole conversation, with a voice for each line, in one file.

Voices

Change, create and translate voices. Voice Changer gives a recording or a talking video a new voice, and Voice Changer Pro gives each speaker in a conversation a voice of their own. Voice Design and Voice Remix create a new voice from a description, and Dubbing translates speech or a whole video into another language.

Music

Make and edit songs. Suno Create Music makes a full song from a prompt or your own lyrics, and Generate Music makes a track from a reference song plus a prompt. The other Suno nodes write lyrics, make covers and mashups, extend a track, replace a section, add vocals or an instrumental, and convert a song to WAV.

Sound Effects

Generate sound from a description. Text to Audio makes sound effects and ambience up to 22 seconds long, including loops that repeat without a gap.

Clean & Separate

Take sound apart. Voice Extractor keeps only the voice and removes noise and music. Audio Separation splits any song into vocals and instrumental or into stems, and Suno Separate does the same for songs made with Suno.

Analyze

Measure a recording without changing it. Silence Detect finds the silent spans to cut, and Audio Sync measures how far apart several recordings of one conversation are. Both run locally, with no AI model and no provider key.

Edit Audio

Cut, join, layer and shape sound. Trim Audio cuts a section, Combine Audio joins clips end to end, and Mix Audio layers tracks with a volume for each. Adjust Volume changes the level and adds fades, and Audio FX adds reverb, telephone, megaphone and echo effects.

Transcribe

Turn speech into text and timings. Transcribe writes down what was said, with word timings for captions, speaker labels and audio events. Forced Alignment times each word of a script you already have.

Which audio node do I need?

You want toUse
Use your own audio file or linkUpload Audio
Use the sound of a YouTube videoReference Audio
Read a script aloud with one voiceText to Speech
Voice a conversation between several speakersText to Dialogue
Change the voice in a recording or a videoVoice Changer
Give each speaker in a conversation a new voiceVoice Changer Pro
Create a new voice from a description, and keep itVoice Design
Try voice ideas quicklyVoice Remix
Translate speech or a video into another languageDubbing
Make a full song from a prompt or lyricsSuno Create Music
Make music from a reference trackGenerate Music
Make a sound effect or an ambienceText to Audio
Remove background noise from speechVoice Extractor
Split any song into vocals, instrumental or stemsAudio Separation
Split a song made with SunoSuno Separate
Find the silences to cut from a recordingSilence Detect
Line up several recordings of one conversationAudio Sync
Cut a section out of a fileTrim Audio
Play clips one after anotherCombine Audio
Layer a voice over musicMix Audio
Change the volume, level the loudness or add fadesAdjust Volume
Put a voice in a room, or make it sound like a phoneAudio FX
Turn speech into text or captionsTranscribe
Time the words of a script you already haveForced Alignment

Nodes that put sound on a video, such as Merge Video & Audio, Extract Audio and Video SFX, are in the Video tab, under Sound for Video. The AI models behind the audio nodes, with their credit prices, are listed in Audio models.

Every audio node

Add Your Own

Speech & Voiceover

Voices

Music

Sound Effects

Clean & Separate

Analyze

Edit Audio

Transcribe

Frequently asked questions

Last updated on

On this page