Skip to content

Choosing a model

The model catalog is curated and moves fast — new models appear in the dropdowns without any change to your graphs, so this page teaches how to choose, not a frozen inventory. The dropdown on the node is always the source of truth for what’s available today: searchable, grouped by provider, with the per-run price right next to the pick.

The model picker on an image node: search box and provider-grouped model list

You want to… Reach for
Compose a new image from instructions, references respected The instruction-following flagships (Nano Banana Pro / 2, GPT Image 2). They read long prompts carefully and handle “the jacket from reference 1 on the person from reference 2” briefs best.
Edit an existing image faithfully — change one thing, keep the rest The dedicated edit variants (…-edit models, the Kontext family, Qwen Edit). They’re anchored to the source image by design; ask for the delta, not a re-description.
Photoreal look development The Seedream family — strong at photographic lighting and texture, with matching edit variants to stay in-family.
Text and lettering in the image Typography-strong models (the Ideogram family). Most other models still mangle words.
A recurring character across many images Character-focused models (character/reference variants in the dropdown) — or better, reference discipline: see Reference images.

Two catalog-wide conventions:

  • -pro / -premium tiers trade cost for quality. Draft on the standard tier, re-run the keeper on the premium one — see Usage & costs.
  • A model that ignores a parameter simply doesn’t show it — if there’s no strength or temperature knob on the card, that model doesn’t take one.
  • The flagship generation family (Seedance 2.5 and siblings) is the default: up to 1080p, native audio, and the full control surface — start/end frames or a multimodal reference kit (see Video prompting for why it’s or).
  • Cost tiers exist here toofast, mini, and offpeak variants render the same kind of shot for less; use them for drafts and blocking, then re-render finals on the flagship.
  • Specialists: human-performance animation models for people and faces, and the swap models that power Character Swap. If the shot is about a person doing something specific, try a specialist before brute-forcing the generalist.
  • Fast tier (Haiku, Flash-class): prompt expansion, batch planning, quick descriptions — most canvas LLM work belongs here.
  • Balanced tier (Sonnet-class, GPT): script drafts, structured briefs, anything a person will read.
  • Heavy tier (Opus-class): the cinematic writing chain with critique enabled, nuanced brand voice, long-context work.

Run the same prompt through two candidates side by side — duplicate the node, switch the model, run both. On a canvas that comparison costs one wire and a few cents, and it answers the question better than any table.