Nodaro Docs
DocumentationNode ReferenceModelsAI Agents (MCP)DevelopersSelf-hostingResearch
Video models

Video models

Every video model you can run in Nodaro, with what each one does and what it costs in credits.

Nodaro runs 62 video models. Each row below links to a page with the model's settings, credit prices and prompt tips.

MiniMax

ModelMakerModesCreditsDetails
Hailuo 02 I2V ProMiniMaxImage to video, Text to video143Hailuo 02 Pro — strong photoreal motion, fixed 5-second clips. Supports end frame.
minimax-h3MiniMaxImage to video, Text to videofrom 230MiniMax Hailuo 3 — premium multimodal tier: first/last frame + image/video/audio references, native audio, 2K (default) or 768P output, 4-15s per-second pricing.
Hailuo 2.3 ProMiniMaxImage to videofrom 130Hailuo 2.3 Pro — newer Hailuo with 768P / 1080P resolutions.
Hailuo 2.3 StandardMiniMaxImage to videofrom 75Cheaper Hailuo 2.3 tier — good baseline quality.
Hailuo 02 StandardMiniMaxImage to video, Text to videofrom 75Hailuo 02 Standard — economical option with end-frame support.

Google

ModelMakerModesCreditsDetails
VEO 3.1 QualityGoogleImage to video, Text to videofrom 930Google VEO 3.1 Quality — premium cinematic video. 4/6/8s clips, optional end frame, native audio. No reference-to-video mode (Fast/Lite only). Flat per-generation pricing across durations.
VEO 3.1 FastGoogleImage to video, Text to videofrom 150VEO 3.1 Fast — cheaper VEO 3.1 tier, 4/6/8s with audio. Good balance for most uses. Flat per-generation pricing across durations.
VEO 3.1 LiteGoogleImage to video, Text to videofrom 75VEO 3.1 Lite — most cost-effective VEO tier for high-volume generation. 4/6/8s with audio, supports first+last frame.
Gemini OmniGoogleImage to video, Text to videofrom 230Google multimodal video with native audio; text/image-to-video + video-edit.
Gemini Omni FlashGoogleImage to video, Text to videofrom 160Google Gemini Omni Flash — faster/cheaper Omni tier: multimodal video with native audio, text/image-to-video + video-edit.
VEO ExtendGoogleVideo extensionfrom 190Extend an existing VEO 3.1 clip by another segment.
VEO 1080p UpscaleGoogleVideo upscaling20Upscale VEO output to 1080p.
VEO 4K UpscaleGoogleVideo upscaling380Upscale VEO output to 4K.

Kuaishou

ModelMakerModesCreditsDetails
Kling 2.6KuaishouImage to video, Text to videofrom 138Kling 2.6 I2V — strong motion realism. 5s/10s, optional native audio.
Kling 2.5 Turbo ProKuaishouImage to video, Text to videofrom 110Faster Kling — good quality at lower cost. Supports end frame.
Kling 3.0KuaishouImage to video, Text to videofrom 270Premium Kling 3.0 — variable 3-15s duration, native audio, 720P/1080P.
Kling 2.1 MasterKuaishouImage to videofrom 400Master tier I2V — strong cinematic quality.
Kling 3 OmniKuaishouImage to videofrom 250Kling 3 Omni — 3-15s, 720p/1080p, end frame + reference images, native audio.
Kling 2.6 Motion TransferKuaishouMotion transferfrom 80Transfer the motion from a driving video onto a still subject. Kling 2.6 base.
Kling 3.0 Motion TransferKuaishouMotion transferfrom 150Premium motion transfer via Kling 3.0.
Kling Avatar StandardKuaishouLip sync280Lip-sync a still portrait to driving audio. Standard quality.
Kling Avatar ProKuaishouLip sync560Premium lip-sync — better mouth shape and timing.

xAI

ModelMakerModesCreditsDetails
Grok Imagine (I2V)xAIImage to videofrom 50Grok image-to-video — stylized motion. Up to 15s.
Grok Imagine Video 1.5xAIImage to videofrom 295Grok Imagine 1.5 image-to-video — 1–15s, 480p/720p, per-second pricing. Requires an input image.

Bytedance

ModelMakerModesCreditsDetails
Seedance 2BytedanceImage to video, Text to videofrom 230Seedance 2 — premium tier with native audio. Per-second pricing by resolution.
Seedance 2 FastBytedanceImage to video, Text to videofrom 180Cheaper / quicker Seedance 2 tier.
Seedance 2 MiniBytedanceImage to video, Text to videofrom 120Budget Seedance 2 tier — 480p/720p only, per-second pricing by resolution.
Seedance 2.5BytedanceImage to video, Text to videofrom 340Seedance 2.5 — up to 30s in one shot, native audio, wide multimodal references. 480p/720p/1080p.
Bytedance Lite I2VBytedanceImage to video, Text to video57Cheapest Bytedance video tier with end-frame support.
Bytedance Pro I2VBytedanceImage to video, Text to video175Pro Bytedance video tier — better quality.
Bytedance Pro Fast I2VBytedanceImage to video90Faster Bytedance Pro variant.
Seedance 2 ExtendBytedanceVideo extensionfrom 260Extend ANY video: generates the continuation (audio included) and trim-stitches it into one seamless clip.
OmniHuman 1.5BytedanceLip syncfrom 1020Premium prompt-directed talking avatar from a still image + audio. 720p / 1080p, up to 60s. People, pets, anime.

Alibaba

ModelMakerModesCreditsDetails
Wan 3.0AlibabaImage to video, Text to videofrom 160Wan 3.0 — multimodal: first/last frame or image/video/audio references, native audio, 2-30s at 480p/720p/1080p.
Wan 3.0 PrimeAlibabaImage to video, Text to videofrom 250Wan 3.0 Prime — Alibaba's high-speed Wan 3.0 tier: same multimodal surface and 2-30s range, faster turnaround at a higher per-second rate.
Wan 2.6 I2VAlibabaImage to videofrom 175Wan 2.6 image-to-video — 5/10/15s at 720p/1080p.
Wan 2.2 TurboAlibabaImage to video, Text to videofrom 100Cheap, fast Wan turbo — 5s. Serves both i2v and t2v under one id.
Wan 2.6AlibabaVideo to video, Text to videofrom 175Wan 2.6 — text-to-video and video-to-video under a single id.
Wan Flash V2VAlibabaVideo to video100Faster Wan V2V variant.
Wan 2.7 VideoEditAlibabaVideo to video320Guided video editing with optional reference image, audio control, and prompt expansion.
Wan 2.7 I2VAlibabaImage to video188Wan 2.7 image-to-video — 2–15s at 720p/1080p, supports start+end frame.
Wan 2.7 T2VAlibabaText to video188Wan 2.7 text-to-video — 2–15s at 720p/1080p.

Lightricks

ModelMakerModesCreditsDetails
LTX 2.3 ProLightricksImage to video, Text to videofrom 40Lightricks LTX 2.3 Pro — text/image/audio→video up to 4K, 6/8/10s, end-frame interpolation.
LTX 2.3 FastLightricksImage to video, Text to videofrom 180Lightricks LTX 2.3 Fast — text/image→video up to 20s at 1080p (6/8/10s at 2K and 4K). No audio input, no extend.

HappyHorse

ModelMakerModesCreditsDetails
HappyHorse 1.1HappyHorseText to video282HappyHorse 1.1 text-to-video — 3–15s at 720p/1080p, 9 aspect ratios incl. 21:9/9:21, per-second pricing.
HappyHorse 1.1 I2VHappyHorseImage to video282HappyHorse 1.1 image-to-video — 3–15s at 720p/1080p, aspect ratio inferred from input image, per-second pricing.
HappyHorse 1.1 Ref2VHappyHorseImage to video282HappyHorse 1.1 reference-to-video — 1–9 reference images, 3–15s at 720p/1080p, per-second pricing.
HappyHorse EditHappyHorseVideo to video350HappyHorse video-edit — video-to-video transformation, up to 60s input, 720p/1080p output.

Runway

ModelMakerModesCreditsDetails
Runway Gen-3RunwayImage to video, Text to video30Runway Gen-3. 5/10s at 720p/1080p.
Runway Aleph V2VRunwayVideo to video350Runway Aleph — video-to-video conversion.
Runway ExtendRunwayVideo extension320Extend a Runway video by another clip.

Topaz

ModelMakerModesCreditsDetails
Topaz Video UpscaleTopazVideo upscaling190High-quality video upscale and enhancement.

InfiniTalk

ModelMakerModesCreditsDetails
InfiniTalkInfiniTalkLip syncfrom 110Audio-driven talking-head from a still image. 480p / 720p.

Sync

ModelMakerModesCreditsDetails
Sync Lipsync v3SyncLip syncfrom 1000Dub existing footage — re-syncs lips to a new audio track. Video input, billed per second.

Volcengine

ModelMakerModesCreditsDetails
Volcengine Lip SyncVolcengineLip syncfrom 300Video-to-video AI dubbing — re-syncs lips to a new vocal track. Multi-speaker (scene detection + speaker ID) in basic mode. Video input, billed per second.

Nodaro

ModelMakerModesCreditsDetails
Video Analysis (Fast — legacy)NodaroVideo analysisfrom 181Legacy fast-tier analysis model (pre-2026-07). Kept so stored raw-model configs keep running and keep pricing under their own identifier; new fast-tier runs use the current fast model.
Video Analysis (Fast)NodaroVideo analysisfrom 205Analyze a video into a structured shot list (scenes, camera, audio) — fast, economy tier. Billed per duration bucket.
Video Analysis (Pro)NodaroVideo analysisfrom 217Analyze a video into a structured shot list (scenes, camera, audio) — higher-fidelity, default tier. Billed per duration bucket.
Video Analysis (Mixed)NodaroVideo analysisfrom 270Our most advanced analysis tier — multiple analysis engines combined into one result for maximum completeness and accuracy. Billed per duration bucket.
Video Analysis (Smart)NodaroVideo analysisfrom 414Highest-accuracy analysis — a hybrid pass that blends a native skeleton read with several donor analysis rolls, then always refines the merged result to find shot boundaries and identify the cast by appearance. Best choice when the shot list will drive regeneration. Billed per duration bucket.
AI AuditNodaroVideo auditfrom 215Re-watches a clip against a wired analysis, applies video-verified corrections under guards, and returns a disclosed report of what changed. Billed per duration bucket.
AI Audit (with analysis run)NodaroVideo auditfrom 396Same video-verified audit as AI Audit, but auto-runs a fast analysis first when none is wired in, then applies corrections under guards and returns a disclosed report. Billed per duration bucket.

Last updated on