How many reference images do AI video models take?
The notes that give a number: Veo 3.1 up to three images, Vidu one to seven, Wan 2.7 up to five images or videos mixed, HappyHorse 1.0 up to nine for generation and five for its edit model. Runway, Sora, Kling, PixVerse and Grok describe references without a count in the notes read. As of 2026-09-25.
| Date | Line | Version | What changed |
|---|---|---|---|
| 30 April 2025 | Runway | Gen-4 | Gen-4 References keeps characters and locations consistent |
| 26 August 2025 | Vidu | Not named in the note | Reference video generation from one to seven images |
| 15 October 2025 | Google Veo | Veo 3.1 | Veo 3.1 preview adds extension, up to three reference images and first-and-last-frame input |
| 16 December 2025 | Wan | Wan 2.6 | Wan 2.6 reference-to-video keeps a person's look and voice, with several characters at once |
| 25 February 2026 | Kling AI | Kling 3.0 | Kling 3.0 and 3.0 Omni reach the API; elements can be built from video |
| 12 March 2026 | Sora | Sora 2 | Character references, clips up to 20 seconds, 1080p on Sora 2 Pro, extensions and batch jobs |
| 3 April 2026 | Wan | Wan 2.7 | Wan 2.7 reference-to-video mixes up to five image or video references and clones a voice timbre |
| 26 April 2026 | HappyHorse | HappyHorse 1.0 | HappyHorse 1.0 reference-to-video takes up to nine reference images |
| 26 April 2026 | HappyHorse | HappyHorse 1.0 | HappyHorse 1.0 video edit changes parts of a clip from instructions and up to five images |
| 26 July 2026 | PixVerse | PixVerse V6 | V6 reference-to-video takes video references in an omni mode |
| 31 July 2026 | Grok Imagine | Grok Imagine 1.5 | Grok Imagine video 1.5 adds reference-to-video with preset voices and native 1080p |
| Undated version row | Seedance | Seedance 2.0 | Seedance 2.0: 4 to 15 seconds, up to 4K, multimodal references, editing and extension |
Inclusion rule. Entries whose note bears directly on the question. Order. By date, oldest first; undated rows last.
A reference image tells the model what a person, object or place looks like so it stays the same across shots. The notes describe it under several names: references at Runway and Veo, elements at Kling, character references at Sora, Fusion at PixVerse. The idea is the same; the limits differ, and most notes do not state them.
Counts rose quickly where they are given, from three in Veo 3.1's October 2025 note to nine in HappyHorse 1.0's in April 2026. Two 2026 notes go further than pictures: Wan 2.7 mixes images with video references, and PixVerse V6 takes video references in an omni mode. Several notes tie a voice to the reference too.
1In the makers' words
Runway, 30 April 2025
Generate consistent characters, locations and more.
Runway, changelog
- Gen-4 ReferencesGen-4 References keeps characters and locations consistent
Vidu, 26 August 2025
Added reference video generation - Supports uploading 1–7 images
Vidu, API update notice
- August 26Reference video generation from one to seven images2025
Google Veo, 15 October 2025
Released Veo 3.1 and 3.1 Fast models in public preview, with new features including: Extending Veo-created videos. Referencing up to three images to generate a video. Providing first and last frame images to generate videos from.
Google, Gemini API changelog
- October 15Veo 3.1 preview adds extension, up to three reference images and first-and-last-frame input2025
Wan, 16 December 2025
supports using a specified person or any object as a reference, precisely maintaining consistency of appearance and voice, and allows multi‑character reference for joint performances.
Alibaba Cloud Model Studio, newly released models
- wan2.6-r2vWan 2.6 reference-to-video keeps a person's look and voice, with several characters at once
Kling AI, 25 February 2026
Supports the creation of elements through video
Kling AI, API updates
- 02/25/2026Kling 3.0 and 3.0 Omni reach the API; elements can be built from video
Sora, 12 March 2026
Expanded the Sora API with reusable character references, longer generations up to 20 seconds, 1080p output for sora-2-pro , video extensions, and Batch API support for POST /v1/videos .
OpenAI, API changelog
- March 12Character references, clips up to 20 seconds, 1080p on Sora 2 Pro, extensions and batch jobs2026
Wan, 3 April 2026
Supports hybrid referencing of up to 5 mixed image/video inputs and audio timbre cloning.
Alibaba Cloud Model Studio, newly released models
- wan2.7-r2vWan 2.7 reference-to-video mixes up to five image or video references and clones a voice timbre
HappyHorse, 26 April 2026
Capable of processing up to 9 reference images, it precisely preserves creative intent to deliver superior performance.
Alibaba Cloud Model Studio, newly released models
- happyhorse-1.0-r2vHappyHorse 1.0 reference-to-video takes up to nine reference images
HappyHorse, 26 April 2026
It allows for local or global editing of video elements using up to 5 reference images, precisely preserving original motion dynamics to achieve superior expressiveness.
Alibaba Cloud Model Studio, newly released models
- happyhorse-1.0-video-editHappyHorse 1.0 video edit changes parts of a clip from instructions and up to five images
PixVerse, 26 July 2026
V6 Reference-to-Video now supports `reference_mode: "omni"` and `video_references`.
PixVerse, API changelogs
- 2026/07/26V6 reference-to-video takes video references in an omni mode
Grok Imagine, 31 July 2026
grok-imagine-video-1.5 now supports text-to-video, image-to-video, and reference-to-video (including optional preset voices), with native 1080p for T2V and I2V.
xAI, API release notes
- July 31Grok Imagine video 1.5 adds reference-to-video with preset voices and native 1080p
Seedance, Undated version row
4K (10-bit color depth)
BytePlus ModelArk, model list
- dreamina-seedance-2-0-260128Seedance 2.0: 4 to 15 seconds, up to 4K, multimodal references, editing and extension
References and consistency · Which AI video models take a first and a last frame? · Which AI video models can extend a clip?
2Notes read
- Runway, changelog, read 2026-09-25
- Vidu, API update notice, read 2026-09-25
- Google, Gemini API changelog, read 2026-09-25
- Alibaba Cloud Model Studio, newly released models, read 2026-09-25
- Kling AI, API updates, read 2026-09-25
- OpenAI, API changelog, read 2026-09-25
- PixVerse, API changelogs, read 2026-09-25
- xAI, API release notes, read 2026-09-25
- BytePlus ModelArk, model list, read 2026-09-25