Pricing

Multi-reference

Multi-reference video, from every model that takes it.

You stop rebuilding a character every shot. The references go in together. Prices are on the cards.

Render from your references

The definition

What is multi-reference AI video?

Multi-reference AI video is a generated clip composed from several images at once. You hand the model a face, a product and a set, and it keeps all of them across the take. Each model publishes its own ceiling on how many it will read, and the cards below carry that number straight from the platform catalog. Models that read one image only, or that read two as the first and last frame of a shot, are a different job and are not on this page.

The roster

Every model that reads several references.

Read the cap and the price. Then open the page.

Seedance 2.5

ByteDance

30s in one take. 30 references in one pass.

References
Up to 30
Max length
30s
Resolutions
480p / 720p

from ⚡85$0.85 at pack rate

Open the model page

Seedance 2.0

ByteDance

Native 4K output, strong prompt adherence.

References
Up to 9
Max length
15s
Resolutions
480p / 720p / 1080p / 4K

from ⚡56$0.56 at pack rate

Open the model page

Seedance 2.0 Fast

ByteDance

Seedance 2.0 quality, faster & cheaper — drafts and volume.

References
Up to 9
Max length
15s
Resolutions
480p / 720p

from ⚡45$0.45 at pack rate

Open the model page

Seedance 2.0 Mini

ByteDance

Compact Seedance 2.0 — budget-friendly clips.

References
Up to 9
Max length
15s
Resolutions
480p / 720p

from ⚡28$0.28 at pack rate

Open the model page

Veo 3.1

Google DeepMind

Sharper prompt adherence than the previous generation.

References
Up to 3
Max length
8s
Resolutions
720p / 1080p

from ⚡326$3.26 at pack rate

Open the model page

Veo 3.1 Fast

Google DeepMind

Veo quality, faster & cheaper.

References
Up to 3
Max length
8s
Resolutions
720p / 1080p

from ⚡123$1.23 at pack rate

Open the model page

Grok Imagine Video

xAI

Expressive, stylized looks — pattern-interrupt creatives.

References
Up to 8
Max length
15s
Resolutions
480p / 720p

from ⚡11$0.11 at pack rate

Open the model page

Gemini Omni Flash

Google

Google any-to-any video — built-in sound, four-rung length ladder.

References
Up to 6
Max length
10s
Resolutions
360p / 720p / 1080p / 4k

from ⚡95$0.95 at pack rate

Open the model page

MiniMax H3

MiniMax

Reads a written brief and reference stills in one pass.

References
Up to 9
Max length
15s
Resolutions
768P / 2K

from ⚡66$0.66 at pack rate

Open the model page

Kling O1

Kuaishou

Multi-reference omni model — motion transfer and clip editing.

References
Up to 7
Max length
10s
Resolutions
720p / 1080p

from ⚡77$0.77 at pack rate

Open the model page

Happy Horse 1.1 (reference-to-video)

Alibaba

Up to 9 references, cited in the prompt in attach order.

References
Up to 9
Max length
15s
Resolutions
720P / 1080P

from ⚡86$0.86 at pack rate

Open the model page

MiniMax H3 Max (image-to-video)

MiniMax

Animate a first frame — a second image lands the closing frame.

References
Up to 2
Max length
15s
Resolutions
480P / 768P

from ⚡51$0.51 at pack rate

Open the model page

MiniMax H3 Max (reference-to-video)

MiniMax

Up to 9 references cited by position in the prompt (Image 1, Video 1…).

References
Up to 9
Max length
15s
Resolutions
480P / 768P

from ⚡8$0.08 at pack rate

Open the model page

MiniMax H3 Max Turbo (image-to-video)

MiniMax

Animate a first frame at Turbo prices; no reference lane here.

References
Up to 2
Max length
15s
Resolutions
480P / 768P

from ⚡26$0.26 at pack rate

Open the model page

Wan 2.7 (image-to-video)

Alibaba

Animate a first frame — or bridge to a last one.

References
Up to 2
Max length
15s
Resolutions
720P / 1080P

from ⚡41$0.41 at pack rate

Open the model page

Runway Aleph 2

Runway

Video-to-video restyling — bring your own clip.

References
Up to 5
Max length
30s

⚡3/s$0.03/s at pack rate

Open the model page

Real renders

Made here. Prompt included.

Clips from the gallery. Every prompt is verbatim.

Three moves

References to a finished shot.

01

Pick the model

Choose from the roster above. The card shows its cap.

02

Attach the set

Drop in the face, the product and the location.

03

Write and render

The exact price sits on the button. Press it once.

One shot, several sources

The hard part of generated video was never the first clip. It was the second one, with the same face in it. Reference images are how that got solved: you hand the model the material instead of describing it, and the take is built from what you gave it. What this page adds is the shortlist. Instead of opening every model page to find out which ones read more than one image, you read one roster that the platform catalog writes.

The cap is per model, and it is the wire's own

Every model here publishes its own ceiling and each card prints that model's number. The same figure gates the composer's attach controls and the spec sheet on the model page, so those surfaces cannot disagree. Length ceilings and resolutions come from the same feed. If a vendor widens the envelope, this page widens with it.

References, not frames

A model that takes two images because a shot has a first frame and a last frame is not doing this job, and carding it here would sell the wrong thing. The roster reads the roles each model publishes for its attachments and keeps the frame-only models off the page. Where a model names no roles at all, its images are references and it is listed.

The price is visible before you spend

Each card carries a floor price in credits and the same figure in dollars at the pay-as-you-go pack rate. Credits sell at several rates and plans buy them cheaper, which is why the dollar figure says which rate it came from. The binding quote is still the one on the Generate button, computed for the exact length and tier you picked.

Questions

Frequently asked questions

The roster on this page is the answer, and it is read from the platform catalog rather than written down. Every model listed accepts more than one reference on a single render, and each card names the vendor behind it, the cap it publishes, the longest clip it will carry and the price of a render. A model whose vendor raises that cap shows the new number here on the next catalog refresh.

It depends on the model, so each card states its own ceiling and no page-wide number is claimed. The figure is the maximum the wire itself accepts, read from the same feed the composer reads before it lets you attach. Go over it and the attach controls stop you in the composer rather than at dispatch, so a full set never costs you a failed render.

No, and this page keeps them apart. A first frame is the opening picture of the shot, and a first-and-last pair is the two ends of one move. A reference is material the model composes from: a face to keep, a product to place, a look to match. Models that publish only frame roles are doing the first job and are not carded here, even where they accept two images.

That is what they are for, and the model page for each one states its own behaviour rather than a shared claim. Several angles of the same face give the model more to hold onto across a take than one photograph does. Read the model page before you plan a shoot around a specific character, since vendors differ on what they keep and what they reinvent.

Free to try. A new account gets a one-time credit grant and no card is asked for. Packs start at $10/1,000⚡. Video costs more than stills, so the grant is there to judge the output rather than to fund a campaign.

Render from your references

Pick a model. Attach the set. Press once.

Render from your references

The price is on the button before you press it.