Audio models
Audio models
Every audio model you can run in Nodaro, with what each one does and what it costs in credits.
Nodaro runs 20 audio models. Each row below links to a page with the model's settings, credit prices and prompt tips.
ElevenLabs
| Model | Maker | Modes | Credits | Details |
|---|---|---|---|---|
| ElevenLabs v3 | ElevenLabs | Text to speech | 30 | Latest ElevenLabs TTS — supports [audio tags] for emotion / pacing. Direct API. |
| ElevenLabs Turbo v2.5 | ElevenLabs | Text to speech | 15 | Fast, cheap ElevenLabs TTS via the direct ElevenLabs API. Good for narration. |
| ElevenLabs Multilingual v2 | ElevenLabs | Text to speech | 30 | Multi-language ElevenLabs TTS via the direct ElevenLabs API. |
| ElevenLabs Dialogue v3 | ElevenLabs | Multi-speaker dialogue | 25 | Multi-speaker dialogue via the direct ElevenLabs API — give it a script, it voices each role (any voice: premade, library, or cloned). |
| ElevenLabs Voice Design | ElevenLabs | Voice design | 50 | Design a synthetic voice from a description (no reference clip needed). |
| ElevenLabs Voice Changer | ElevenLabs | Voice changer | 40 | Speech-to-speech: convert one voice to another while preserving prosody. |
| ElevenLabs STT | ElevenLabs | Speech to text | 22 | Speech-to-text with WORD-level timestamps (always on), speaker diarization and audio-event tags. The engine to use when the transcript feeds captions. |
| ElevenLabs Voice Isolation | ElevenLabs | Voice isolation | 74 | Strip background noise / music from a vocal track. |
| ElevenLabs Dubbing | ElevenLabs | Dubbing | 40 | Translate + dub audio or a whole video into a new language — video in, dubbed video out. Async. |
| ElevenLabs Dubbing v2 | ElevenLabs | Dubbing | 1100 | Translate + dub audio or a whole video into a new language — video in, dubbed video out. Async. |
| ElevenLabs Forced Alignment | ElevenLabs | Forced alignment | 30 | Align an existing transcript to audio with word-level timestamps. |
| ElevenLabs Sound Effects | ElevenLabs | Sound effects | 3 | Generate short sound effects from a text prompt. |
OpenAI
| Model | Maker | Modes | Credits | Details |
|---|---|---|---|---|
| Incredibly Fast Whisper | OpenAI | Speech to text | 40 | Fast Whisper speech-to-text. Returns WORD-level timestamps when asked, so its transcript can feed captions. |
| Whisper | OpenAI | Speech to text | 40 | Whisper speech-to-text — PHRASE-level segments only, NO word timestamps. Fine for a transcript or a static subtitle; not for word-timed (kinetic) captions. |
Suno
| Model | Maker | Modes | Credits | Details |
|---|---|---|---|---|
| Suno v4 | Suno | Music | 30 | Suno v4 music generation — full songs with vocals, multiple genres. |
| Suno v5 | Suno | Music | 30 | Suno v5 — better vocal quality than v4, more genres. Same price. |
| Suno v5.5 | Suno | Music | 30 | Suno v5.5 — improved audio quality and expressiveness over v5. |
| Suno V6 | Suno | Music | 30 | Suno V6 — greater musical expression with more natural vocals and richer details. The flagship and the default. |
| Suno V6 Wild | Suno | Music | 30 | Suno V6 Wild — pushes creative boundaries for bolder, more distinctive musical expression; more varied, less predictable results. |
| Suno V6 Mini | Suno | Music | 30 | Suno V6 Mini — lightweight and fast, balancing quality and speed for effortless creation. |
Last updated on