Video production — plan deep, generate elsewhere

The most important and most involved post type. GoUltra takes you as far as you want: most videos stop at the script; when you choose to go deeper it breaks the script into clips → shots, plans & generates the asset images, and writes the model-tuned prompt for each clip. GoUltra never generates the video itself — you take the prompts + assets to your chosen tool (Higgsfield / Google Flow) and bring the cut back. Demo brand: Fidelis Global · a 30s client-proof reel.

Always
Script
Shape the script in conversation. The default stop — hand off a brief, done.
Opt-in
Clips → Shots
Pick the video model; AI breaks the script into clips (within the model's max), each into shots.
Opt-in
Assets
A reusable cast + locations library; generate the images each clip/shot needs (or upload what you have).
External
Hand off
Per-clip prompts + assets → your tool → bring the final cut back to schedule.
Progressive depth. A normal post stops at the script. Going into clips & assets is opt-in — and everything we make (clips, shots, asset images, prompts) is to feed the external generator, never to replace it.
The generation profile drives everything — pick one per video

The model you'll generate on isn't where we make the video — it's the reference that sets each clip's max length, its approach, its shot style, and which assets it needs. Specs verified 2026.

Seedance 2.0
Higgsfield · ByteDance
Approach ingredients / references → video
Clip length up to 15s (aim high — fewer clips, faster)
Multi-shot one-gen multi-camera
Inputs up to 9 image refs + first/last frame
Best for: character-driven, multi-shot storytelling with consistent identity.
Veo 3.1
Google Flow
Approach frames → video (first, or first + last)
Clip length 8s default (4 / 6 also; 10s on Fast)
Multi-shot via Extend (chain)
Inputs first frame [+ last frame]
Best for: exact text & infographics — text is baked into the keyframe and stays crisp.
Gemini Omni Flash
Google Flow
Approach ingredients / avatar → video
Clip length up to 10s
Multi-shot scene-memory edits · 1-gen TBD
Inputs character / avatar + audio refs
Best for: avatar/character continuity & conversational iterative edits.
Planner picks the fit. The AI reads your script and recommends the profile — Veo 3.1 for this text/infographic-led reel — which you confirm. The registry is data-driven, so new models drop in as config.
The flow
1 Script default stop Shape the script in conversation (hook → beats → settings). Hand off a brief and you're done — or take the opt-in step into clips & assets. view → 2 Profile + clip breakdown opt-in Pick the generation profile (planner-recommended), then AI breaks the script into clips — each within the model's max duration, aiming high to minimize clip count — fully editable (merge / split / re-time). view → 3 Clip — shots + assets + prompt the heart A clip → its shots (single, or first→last-frame for Veo); the asset checklist each shot needs; and the model-tuned video prompt to paste into the external tool. Chat-refine on any prompt or asset. view → 4 Asset library — cast · locations · generate A reusable video-level library (characters, locations, products, logo) assigned across clips/shots. Per asset: already-have / upload / generate in-app / get-a-prompt — plus composite ("put this character in this scene → a frame"). view → 5 Hand off + bring the cut back Every clip with its prompt + assigned assets, ready to copy into Higgsfield / Google Flow (or a Claude project). Generate there, then drop the finished video back here to schedule. view →
What GoUltra does — and doesn't
  • Does: finalize the script · break it into model-sized clips → shots · run a reusable asset library · generate the asset images (in-app image engine) · write the model-tuned per-clip prompt · package the hand-off · receive the final cut.
  • Doesn't: generate the video. That always happens on your chosen external tool — the profile is just the reference we plan against.
  • Reuses: the existing external-video groundwork (Veo/Omni clip plan, per-clip resources, the brand-asset pool, the 5-part prompt composer) — extended with Seedance 2.0, the Shots level, the asset library, and composite generation.