Audio nodes
Every audio node in Nodaro, from text to speech, voice changing and dubbing to Suno music, sound effects, stem separation, editing, sync and transcription.
The Audio tab of the node picker holds every node that makes, changes, cleans, edits or reads sound. It covers speech and voiceover, voice changing and dubbing, music, sound effects, stem separation, audio editing, recording sync and transcription. Press Tab on the canvas and open the Audio tab to see the nodes, grouped in nine families.
The nine families
Add Your Own
Bring your own sound into a workflow. Upload Audio takes a file or a direct link and lets you trim it before upload. Reference Audio also extracts the sound track of a YouTube video.
Speech & Voiceover
Turn text into speech with ElevenLabs voices. Text to Speech reads a text with one voice, in up to 46 languages, with audio tags for emotion. Text to Dialogue voices a whole conversation, with a voice for each line, in one file.
Voices
Change, create and translate voices. Voice Changer gives a recording or a talking video a new voice, and Voice Changer Pro gives each speaker in a conversation a voice of their own. Voice Design and Voice Remix create a new voice from a description, and Dubbing translates speech or a whole video into another language.
Music
Make and edit songs. Suno Create Music makes a full song from a prompt or your own lyrics, and Generate Music makes a track from a reference song plus a prompt. The other Suno nodes write lyrics, make covers and mashups, extend a track, replace a section, add vocals or an instrumental, and convert a song to WAV.
Sound Effects
Generate sound from a description. Text to Audio makes sound effects and ambience up to 22 seconds long, including loops that repeat without a gap.
Clean & Separate
Take sound apart. Voice Extractor keeps only the voice and removes noise and music. Audio Separation splits any song into vocals and instrumental or into stems, and Suno Separate does the same for songs made with Suno.
Analyze
Measure a recording without changing it. Silence Detect finds the silent spans to cut, and Audio Sync measures how far apart several recordings of one conversation are. Both run locally, with no AI model and no provider key.
Edit Audio
Cut, join, layer and shape sound. Trim Audio cuts a section, Combine Audio joins clips end to end, and Mix Audio layers tracks with a volume for each. Adjust Volume changes the level and adds fades, and Audio FX adds reverb, telephone, megaphone and echo effects.
Transcribe
Turn speech into text and timings. Transcribe writes down what was said, with word timings for captions, speaker labels and audio events. Forced Alignment times each word of a script you already have.
Which audio node do I need?
| You want to | Use |
|---|---|
| Use your own audio file or link | Upload Audio |
| Use the sound of a YouTube video | Reference Audio |
| Read a script aloud with one voice | Text to Speech |
| Voice a conversation between several speakers | Text to Dialogue |
| Change the voice in a recording or a video | Voice Changer |
| Give each speaker in a conversation a new voice | Voice Changer Pro |
| Create a new voice from a description, and keep it | Voice Design |
| Try voice ideas quickly | Voice Remix |
| Translate speech or a video into another language | Dubbing |
| Make a full song from a prompt or lyrics | Suno Create Music |
| Make music from a reference track | Generate Music |
| Make a sound effect or an ambience | Text to Audio |
| Remove background noise from speech | Voice Extractor |
| Split any song into vocals, instrumental or stems | Audio Separation |
| Split a song made with Suno | Suno Separate |
| Find the silences to cut from a recording | Silence Detect |
| Line up several recordings of one conversation | Audio Sync |
| Cut a section out of a file | Trim Audio |
| Play clips one after another | Combine Audio |
| Layer a voice over music | Mix Audio |
| Change the volume, level the loudness or add fades | Adjust Volume |
| Put a voice in a room, or make it sound like a phone | Audio FX |
| Turn speech into text or captions | Transcribe |
| Time the words of a script you already have | Forced Alignment |
Nodes that put sound on a video, such as Merge Video & Audio, Extract Audio and Video SFX, are in the Video tab, under Sound for Video. The AI models behind the audio nodes, with their credit prices, are listed in Audio models.
Every audio node
Add Your Own
- Upload AudioAdd your own audio to a workflow. Upload an MP3, WAV, M4A, AAC, FLAC or OGG file or paste a link, trim it before upload, and wire it to any audio node.
- Reference AudioBring audio into a workflow from a YouTube video, an uploaded file or a direct link. Reference Audio extracts the sound track and lets you preview it first.
Speech & Voiceover
- Text to SpeechTurn text into natural speech with ElevenLabs v3, Turbo v2.5 or Multilingual v2. Choose a voice, add emotion with audio tags, and speak up to 46 languages.
- Text to DialogueVoice a whole conversation in one audio file. Give each line its own voice, add audio tags for emotion, or fill the lines from a script. ElevenLabs Dialogue v3.
Voices
- Voice ChangerReplace the voice in a recording or a talking video with another voice, and keep the original emotion, pacing and timing. Speech-to-speech with ElevenLabs.
- Voice Changer ProDetect each speaker in a recording or talking video and give each one a new voice, with the original emotion and timing kept. Priced per minute of speech.
- Voice DesignCreate a new voice from a written description, with control over model, loudness, guidance and seed. Get an audio preview and a voice ID you can keep and reuse.
- Voice RemixDescribe a voice in plain words and hear it speak your preview text. The quick way to explore voice ideas, with no reference recording and no settings to tune.
- DubbingTranslate spoken audio or a whole video into another language and keep each speaker's own voice. Dub a file or a YouTube link, from 40 credits per minute.
Music
- Suno Create MusicMake full songs with Suno V6 and earlier Suno versions. Write lyrics with metatags, set the style and the vocals, or create an instrumental track.
- Generate MusicCreate original music with MiniMax Music from a reference song, voice or instrumental plus a text prompt, with optional lyrics, genre and mood.
- Suno LyricsWrite complete song lyrics from a short description with Suno. The lyrics come with section tags and a suggested title, ready for Suno Create Music.
- Suno CoverMake a cover version of a song with Suno. Change the genre, the lyrics or the vocal gender, sing it in your own voice, or turn it into an instrumental.
- Suno ExtendContinue a Suno song from the second you choose, with a new prompt and style. Add a bridge, an outro or a verse, or chain extensions into a long track.
- Suno MashupBlend two audio tracks into one mashup with Suno. Turn on custom mode to set the title, the style tags, the styles to avoid and the vocal gender.
- Suno Replace SectionRegenerate a 6 to 60 second section of a Suno song with new lyrics and style tags, such as a weak verse or chorus, and keep the rest of the track intact.
- Suno Add VocalsAdd AI-generated vocals to an instrumental track made by a Suno node. Suno writes the lyrics and the melody to fit the music, on Suno V6 by default.
- Suno Add InstrumentalAdd an AI-generated instrumental arrangement to a vocal track made by a Suno node. Suno matches the genre, tempo and key of the vocals automatically.
- Suno Upload ExtendContinue any audio with Suno, not only Suno songs. Extend an upload, a voiceover or another AI track from the second you choose, in the style you set.
- Suno Style BoostTurn a short music style, such as lo-fi chill, into richer Suno style tags, and send the result to the Style field of Suno Create Music for your song.
- Suno Convert WAVConvert a song made by a Suno node from compressed MP3 to lossless WAV, for mixing, mastering, archiving or tools that do not accept MP3 files.
Clean & Separate
- Voice ExtractorRemove background noise, music and other sounds from a recording and keep only the voice. Clean audio before transcription, lip sync, dubbing or a voice change.
- Audio SeparationSplit any audio into vocals and instrumental, or into stems such as drums, bass, guitar and piano, with the Demucs model. Works on songs from any source.
- Suno SeparateSplit a song made with Suno into a vocal and an instrumental track, or into up to 12 stems such as drums, bass, guitar and piano, straight from a Suno node.
Analyze
- Silence DetectFind the silent spans in a recording or a video and get them as time ranges, ready to cut dead air in an edit. A flat 10 credits per run, with no AI model.
- Audio SyncMeasure how far apart 2 to 6 recordings of one conversation are, from their sound, so a multicam edit lines up without typing offsets. Also reports clock drift.
Edit Audio
- Trim AudioCut a section out of an audio file by start and end time, and save it as MP3, WAV or AAC. Keep only the part of a recording or song that the next node needs.
- Combine AudioJoin audio clips end to end into one file, in the order you choose, and trim each clip first. For narration, podcasts and dialogue made in several parts.
- Mix AudioLayer several audio tracks into one file, with a volume from 0 to 200% for each. Put a voiceover over music, or effects over an ambience, before adding video.
- Adjust VolumeChange the volume of an audio track or of a video's sound, level it with Normalize, and add a fade-in and a fade-out of up to 10 seconds each.
- Audio FXPut a voice in a room, a car or a church with a reverb, or add a telephone, megaphone or echo effect. A flat 20 credits per run, plus custom delay and EQ.
Transcribe
- TranscribeTurn the speech in audio or video into text with ElevenLabs STT or Whisper. Get word timings for captions, speaker labels, and tags for music and laughter.
- Forced AlignmentTime every word of a known transcript against its audio and get each word's start and end as data, for karaoke highlights, timed graphics and captions.
Frequently asked questions
Related
Text to Speech
Voice Changer
Suno Create Music
Transcribe
Audio models
Last updated on
AI Audit
Re-watch a video against its Video Analysis, apply only the fixes the footage confirms, and get a report of every fix, declined change and watch item.
Upload Audio
Add your own audio to a workflow. Upload an MP3, WAV, M4A, AAC, FLAC or OGG file or paste a link, trim it before upload, and wire it to any audio node.