Text Models

Text Models are the AI language models available in the Text Node's generation panel. They take your prompt — along with any upstream text, images, or videos — and return a block of text.


Available models

ModelImage inputVideo input
Gemini 3-Flash
Gemini 3.1-Pro
GPT 5.4

The default model is Gemini 3.1-Pro.


Choosing a model

Gemini 3-Flash is the fastest option. Use it for straightforward text tasks — scene descriptions, character bios, simple rewrites — where speed matters more than nuanced output.

Gemini 3.1-Pro delivers higher quality output and handles complex instructions, long upstream contexts, and subtle creative direction more reliably. It also supports video input, making it the right choice when upstream Video Nodes are part of your workflow.

GPT 5.4 is an alternative high-quality option with a slightly different writing style and reasoning approach. Note that it does not accept video input — if any Video Node is connected upstream, those references will not be passed to GPT 5.4.

⚠️

Note: If a generation fails due to too much upstream content, try reducing the number of connected upstream nodes or switching to a different model.


Did this page help you?