Generate Script
Write a multi-scene video script with AI. Each scene gets a description, action, mood, length, camera, dialogue and a ready-to-use image prompt.
The Generate Script node writes a structured, multi-scene video script from a short prompt. Each scene comes with a visual description, the action, a mood, a suggested length and an image prompt ready for Generate Image, plus camera directions, dialogue, location details, a music mood and sound effects when they apply. Seven outputs send each part of the script to the node that needs it.
When to use it
- You want a storyboard for a video: one image prompt per scene, then one image and one clip per scene.
- You need a structured script for an explainer, an ad, a product demo or a social video.
- You want dialogue lines, music moods and sound effects planned together with the pictures.
- You want to plan the length and the pacing of a video before you spend credits on generation.
For a complete film made for you from one prompt, use Story → Video.
Quick start
Add the node
Press Tab on the canvas and choose Video › Story & Script › Generate Script.
Describe the video
Add a Text node, type the topic, the story or the concept in it, and connect it to the Prompt input of Generate Script. Be specific: "30-second product ad for a fitness app showing a morning routine" works better than "make a video".
Set the shape of the script
Open the settings panel. Set Number of Scenes, Target Length (seconds) and Structure, and optionally a Tone and a Style Guide.
Run it and connect the outputs
Click Run. The script appears on the node. Connect its outputs to the nodes that make the pictures, the voices and the music.
Inputs and outputs
The input, Prompt, takes the topic, story or concept from a text node, such as Text or Prompt. Generate Script has no prompt box of its own, so connect one.
The seven outputs each carry one part of the active script:
| Output | What it carries | Connect it to |
|---|---|---|
| Scenes | The script as text: the first scene's image prompt. | A text input. |
| Images | The image prompt of every scene. | The Prompt input of Generate Image. With two or more scenes, Generate Image runs once per scene. |
| Dialogue | Every dialogue line of the script, with speaker and emotion. | Text to Dialogue. Click Fill N Lines from Script in its settings to copy the lines. If Text to Dialogue has no lines of its own, it uses the script's lines when it runs. |
| Music | The music moods of all scenes, each once, joined into one text. | The prompt of Generate Music. |
| SFX | The sound effects of all scenes, joined into one text. | The prompt of Text to Audio. |
| Characters | The characters of the script, each once, with name, description, mood, action and position. | Nodes that read script characters. |
| Locations | The locations of the script, each once, with name, description, time of day, weather and lighting. | Nodes that read script locations. |
Settings
| Setting | What it does |
|---|---|
| AI Model | The text model that writes the script. The default is Gemini 3.6 Flash. Search by name, maker or tier. |
| Reasoning Effort | How much the model thinks before it writes, on models that support it. |
| Advanced mode | Gemini models only. Runs the model with the maker's own controls, so that Temperature, Max Tokens and the full reasoning range apply. Those controls appear once it is on. It costs one credit tier more. On other models, the switch is disabled with the reason. |
| Number of Scenes | How many scenes to write, from 1 to 20. The default is 5. |
| Structure | Freeform (the default) lets the AI choose the pacing. 8-Step Story follows a classic narrative arc. Custom is also available. |
| Style Guide | Visual and narrative directions for every scene, such as a color palette, an era or a visual reference. |
| Tone | The emotional register of the whole script, for example whimsical, dramatic or educational. |
| Target Length (seconds) | The total length the script plans for, from 10 to 600 seconds. The default is 60. The AI spreads it across the scenes, so single scenes can vary. |
| Pre & post text | Text added before and after the prompt, not the style guide, when the node runs. It is hidden from people who use your workflow as an app. See Prompt pre and post text. |
Set these values in the node itself. The Generation Settings nodes, such as Scene Count and Style Guide, cannot pass their values to Generate Script.
What each scene contains
| Part | What it is |
|---|---|
| Scene number and name | The position and a short name of the scene. |
| Visual description | A detailed description of what is seen. |
| Action | What happens in the scene. |
| Mood | One or more emotional tones. |
| Duration | The suggested length in seconds. |
| Image prompt | A prompt ready for Generate Image. |
| Characters | Each character's name, description, mood, action and position. |
| Dialogue | Spoken lines, each with a speaker, the text and an emotion. |
| Location | The name, a description, the time of day, the weather and the lighting. |
| Cinematography | The shot type, the camera angle and the camera movement. |
| Music mood and sound effects | Suggestions for the soundtrack. |
Edit the script
After a run, the settings panel shows the generated script: its Title, the number of scenes and the total length, and one section per scene. Open a scene to edit its Visual Description, Action, Mood, Duration (s) and Image Prompt (for Generate Image). Your edits are what the outputs send. Copy Prompts copies the image prompts of all scenes.
Every run is kept in the node's result strip, and the outputs follow the result you select.
Once the script exists, Expand Storyboard on the node creates the workflow nodes for every scene. Choose Scene Nodes, which creates one scene node per scene, or Pipeline Nodes, which creates separate Generate Image, video, Text to Speech and Merge Video & Audio nodes per scene. You can add a Combine Videos node at the end and start the image generation right away. The dialog shows the estimated credits before you create the nodes.
Presets
Open Presets on the node for format presets that set the tone, the number of scenes and the target length. Type your topic in the prompt.
| Folder | Presets |
|---|---|
| By Format | YouTube Short, Explainer, Ad Spot, Product Demo, Listicle, UGC Ad (a selfie-style Hook, Show, Proof and Opinion) |
| Long-Form & Narrative | Podcast Outline, Trailer Narration, Story Beats |
Credits
A run costs 10, 20 or 30 credits, depending on the tier of the AI model: Economy, Standard or Premium. The default model, Gemini 3.6 Flash, is an Economy model. Advanced mode costs one tier more, and the cost badge on the node updates when you turn it on. See Choosing a model for the tiers.
Tips
- Be specific in the prompt. Name the product, the audience, the length and the format.
- Use Tone for consistency. It applies to every scene, so the whole script keeps one emotional register.
- Put the look in the Style Guide. A palette, an era or a visual reference keeps the image prompts consistent.
- Pick the structure by content. Use 8-Step Story for stories and ads with an arc, and Freeform for informational or documentary videos.
- Treat the length as a plan. Target Length guides the AI, but each scene's length can vary with its content.
Frequently asked questions
Related
Generate Image
Generate Video
Text to Dialogue
Story → Video
Generate Music
Last updated on