Skip to main content
Orisu connects to a range of hosted video models. You pick the model per node in its config panel, so the same node can run on different families depending on what you need. This page covers which models exist and what they’re good for. For what each setting does, see Model settings.

Generation

These families power the Generate video node, turning a text prompt and/or a first-frame image into video. Capabilities vary by family: end frames, reference images and videos, audio input, and subject references appear only on models that support them. Most families now expose their quality tiers directly on the node instead of as separate models: Veo3 and Veo 3.1 offer Standard / Fast (Veo 3.1 adds Lite), Seedance 2.0 offers Standard / Fast / Mini, Kling 3.0 and Kling O3 offer Standard / Pro / 4K, and Kling 2.1 offers Standard / Pro / Master — the credit estimate updates as you switch. Audio on/off is a setting (generate_audio) on models that support it, not a separate model.
Some video models generate audio and some don’t. Families like Veo, Kling, Seedance, and Gemini Omni support sound; models that do expose a generate-audio toggle in the node’s settings (audio usually adds to the per-second price). Turn it off when you plan to add your own soundtrack.

Available generation variants

Veo (Google)

Veo 3.1 (Standard / Fast / Lite), Veo3 (Standard / Fast)

Gemini Omni (Google)

Gemini Omni Flash — text, image, or reference images to video with audio

Sora (OpenAI)

Sora 2

Kling (Kling)

Kling 3.0 (Standard / Pro / 4K), O3 (Standard / Pro / 4K), 2.6 Pro, 2.5 Turbo, 2.1 (Standard / Pro / Master), O1, Elements (Standard / Pro), Motion Control

Hailuo (MiniMax)

Hailuo 02 Pro, 2.3 Standard, 2.3 Fast

Runway (Runway)

Runway Gen 4

Seedance (ByteDance)

Seedance Pro, Pro Fast, 1.5 Pro, 2.0 (Standard / Fast / Mini)

Grok Imagine Video (xAI)

Grok Imagine Video

Hedra / Vidu

Hedra Character 3, Vidu

Editing

The Edit video node rewrites existing footage from text instructions. It runs on:
  • Gemini Omni Flash — conversational edits (object replacement, style changes) without regenerating the whole clip
  • Kling
  • Runway Gen 4 Aleph
  • Grok Imagine Edit

Lip sync

The Generate lipsync node drives a face from an audio track. It runs on:
  • Sync Lipsync 2.0
  • Kling
  • Veed Fabric
  • Hedra

Upscale

The Upscale video node increases resolution. It runs on:
  • FlashVSR
Match the model to the inputs you have. Generation models accept prompts and first frames; editing models need an existing video; lip sync needs a face plus audio; upscaling needs a finished clip.