Video models
Volcengine Lip Sync
Volcengine Lip Sync is a Volcengine video model in Nodaro for lip sync, from 300 credits per run.
Video-to-video AI dubbing — re-syncs lips to a new vocal track. Multi-speaker (scene detection + speaker ID) in basic mode. Video input, billed per second.
Credits
| Tier | Credits |
|---|---|
| 5-min ceiling (no duration given) | 6000 |
| Volcengine Lip Sync | 300 |
| Volcengine Lip Sync | 600 |
| Volcengine Lip Sync | 1200 |
| Volcengine Lip Sync | 2400 |
| 5-min ceiling | 6000 |
Use Volcengine Lip Sync in Nodaro
Add one of these nodes to your canvas, open its settings, and choose Volcengine Lip Sync as the model:
Similar models
| Model | Maker | Modes | Credits | Details |
|---|---|---|---|---|
| Hailuo 02 I2V Pro | MiniMax | Image to video, Text to video | 143 | Hailuo 02 Pro — strong photoreal motion, fixed 5-second clips. Supports end frame. |
| minimax-h3 | MiniMax | Image to video, Text to video | from 230 | MiniMax Hailuo 3 — premium multimodal tier: first/last frame + image/video/audio references, native audio, 2K (default) or 768P output, 4-15s per-second pricing. |
| Hailuo 2.3 Pro | MiniMax | Image to video | from 130 | Hailuo 2.3 Pro — newer Hailuo with 768P / 1080P resolutions. |
| Hailuo 2.3 Standard | MiniMax | Image to video | from 75 | Cheaper Hailuo 2.3 tier — good baseline quality. |
| Hailuo 02 Standard | MiniMax | Image to video, Text to video | from 75 | Hailuo 02 Standard — economical option with end-frame support. |
| VEO 3.1 Quality | Image to video, Text to video | from 930 | Google VEO 3.1 Quality — premium cinematic video. 4/6/8s clips, optional end frame, native audio. No reference-to-video mode (Fast/Lite only). Flat per-generation pricing across durations. |
Frequently asked questions
Last updated on