
This Seedance 2.5 prompt guide covers the official @ reference syntax: 7 modality combos, 50-reference specs, ratio rules, and real ByteDance prompts.
You upload three character images and a reference video into Seedance 2.5, write "reference @image1," hit generate — and the model puts the wrong face on the wrong body. That failure is almost never the model's fault: it's a syntax problem, and ByteDance's own practice guide spells out how to avoid it. This is the first systematic tutorial on the official @ reference syntax — how to bind up to 50 reference files (30 images, 10 videos, 10 audio clips) to specific roles so the output matches what's in your head.
Everything here is based on ByteDance's official Seedance 2.5 enterprise practice guide, cross-checked hands-on in Vogoo's studio — every spec and example prompt comes from that guide or ByteDance's public announcements. Seedance 2.5 officially launched on July 31, 2026, and the reference system is the biggest jump over 2.0: from 9 images and 3 short clips to a 50-file multimodal stack.
The official guide's core prompt formula is simple:
Prompt = generation instructions (text) + reference materials (@Video N @Image N @Audio N)The @ symbol "feeds" a material into the generation — @Image 1, @Video 1, @Audio 1 — and locks the corresponding visual or audio information.
Then comes the rule the guide states explicitly, and the one most people break. Paraphrased from the manual: every time you reference a material with @, immediately follow it with one sentence declaring how that material should be used. With multiple characters or multiple clips, you must spell out the mapping explicitly — otherwise the model gets confused. The guide even names the anti-pattern: a dangling "reference @Image 1" with no head and no tail.
Before — the dangling reference (what the manual warns against):
Reference @Image 1, reference @Image 2, generate an ad videoThe model has no idea whether @image1 is the actor, the product, the scene, or the color grade.
After — every @ carries a usage declaration (official example):
Use the female lead in @Image 1 as the subject, and generate a video
of her strolling through city streetsOne material, one job, stated in the same breath. Scale that discipline up and Seedance 2.5 keeps every face, prop, and scene where it belongs — the wuxia case below does exactly that with 32 references.
The practice guide breaks reference generation into sub-capabilities, each with its own declaration pattern. These official one-liners are worth keeping as templates:
| Reference job | Official prompt pattern (translated) |
|---|---|
| Subject | Use the female lead in @image1 as the subject; generate her strolling through city streets |
| Motion / camera | Reference the camera moves and running rhythm of @video1; take the subject from @image1; regenerate a product ad |
| White-model render | Render from the motion path of the white model in @video1; subject from @image1; scene from @image2 |
| Style | Match the cold-blue night grade and grain of @video1; lead character from @image1 |
| Audio | Match the musical mood of @audio1; visual subject from @image1; generate a rhythm-synced short |
| Storyboard grid | Follow the four-panel storyboard order in @image1; lead from @image2; advance panel by panel |
| Keyframes | @image1 as first frame, @image2 as last frame; fill the in-between transition |
| One-click film | Chain @video1, @video2 and @image1 into a finished short with titles and transitions |
Note the structure repeated in every row: @material + role verb + what it controls. That is the whole syntax — there is no markup beyond @; the binding lives in the sentence you attach to it.
Video editing and extension use a variant formula — edit instructions + @Video 1 (the source clip) + optional @Image N / @Audio N — where the model auto-locks aspect ratio and duration to the source. If you're coming from plain text-to-video or single-image image-to-video workflows, @ referencing is the layer that turns them into a controllable reference-to-video pipeline.
Seedance 2.5 accepts references in any of seven combinations — audio-only being new in 2.5:
| # | Combination | New in 2.5? | Typical use |
|---|---|---|---|
| 1 | Images only | — | Character/product/scene lock, storyboards, keyframes |
| 2 | Video only | — | Motion, camera language, style, white-model transfer |
| 3 | Audio only | Yes — new | BGM or voice drives pacing, beat-syncing, lip-sync |
| 4 | Images + video | — | Subject from stills, movement from a reference clip |
| 5 | Audio + video | — | Motion reference cut to a soundtrack |
| 6 | Images + audio | — | Music video from character art and a track |
| 7 | Images + audio + video | — | Full-stack productions (the flagship demos below) |
Audio-only input is the combination to try first if you make music-driven content — a single track can drive rhythm, cuts, and mouth movement with no visual reference at all. It pairs naturally with an AI music video generator workflow, and voice-driven cases overlap with AI lip sync territory.
The hard limits, straight from the practice guide:
| Modality | Count | Per-file size | Formats | Duration rules |
|---|---|---|---|---|
| Images | 0-30 | See the guide's upload limits | Standard image formats | — |
| Videos | 0-10 | ≤ 200MB, 480p-4K | mp4, mov | 2-30s per clip, all clips total ≤ 30s |
| Audio | 0-10 | ≤ 15MB | mp3, wav | 2-30s per clip, all clips total ≤ 30s |
Two details people miss: video and audio duration budgets are counted separately — up to 30 seconds of reference video and 30 seconds of reference audio in the same job. For context, Seedance 2.0 allowed 9 images and 3 clips each of video/audio with a 15-second total; 2.5 raises that to 30 images, 10 clips each, and 30-second totals.
Fifty slots does not mean you should fill fifty slots. The guide dedicates a section to what its team found through large-scale evaluation:
Want to test these rules without setting up anything? You can try Seedance 2.5 free on Vogoo — 30-second one-shot generation, up to 4K, native synced audio, and up to 50 references in one job, right in the browser.
These are ByteDance's own demos from the practice guide, with the actual prompts where the guide publishes them.
One reference video supplies the creative concept; four images lock the assets. The declaration for each is one clause long:
Reference the creative concept of @Video 1; faithfully reproduce the full
house-assembly process using @Image 1 @Image 2 @Image 3 @Image 4.Official Seedance 2.5 demo — ByteDance / Volcano Engine practice guide; prompt translated from the official guide.
Notice the structure: not just the reference images' appearance, but the build logic from the reference video is carried over.
The guide's pomegranate showcase demonstrates why multimodal referencing exists: one subject travels through multiple scenes with form and style locked, because the reference stack — images, video, audio together — defines it once for the whole generation.
Official Seedance 2.5 demo — ByteDance / Volcano Engine practice guide.
This is the stress test: 32 references — character images, prop images, scene images, sound-effect clips, and an effects reference video — bound segment by segment with timestamps. An excerpt of the official prompt:
[0-3s] Close-up shot: the camera focuses on the hands of @Image 1 plucking
the strings of the guqin @Image 2. On the left, a bronze incense burner
@Image 3 trails wisps of smoke. In the background, the blurred figure of a
white-robed man @Image 4 is playing @Image 5. Sound effects reference
@Audio 1. The scene is set on a small boat @Image 7 on the water @Image 6,
with the distant mountains referencing @Image 8.
[3-5s] The white-robed man @Image 4 stands at the bow... the woman from
@Image 9 attacks wielding @Image 10; sound effects reference @Audio 2.
[5-10s] The white-robed man and the red-clad woman clash fiercely in the air
above the lake; sound effects reference @Audio 3. The red-clad woman holds a
red umbrella, the man holds the fan @Image 11. The water below erupts in
huge ring-shaped waves; sound effects reference @Audio 4, VFX reference
@Video 1.Official Seedance 2.5 demo — ByteDance / Volcano Engine practice guide. Prompt excerpt translated into English from the guide's original.
All 32 references obey the rule from the top of this guide: @ plus an immediate usage declaration — which is what keeps two dueling leads, their weapons, and four sound effects from bleeding into each other.
The guide also shows the syntax in English — the prompt language doesn't change the @ mechanics (Seedance 2.5 natively supports 10-plus languages). The street-dance ensemble prompt opens:
Create a 30-second video with realistic cinematic texture, 16:9 aspect ratio,
documentary street photography style, warm and everyday lifelike atmosphere.
Street scene environment references @image1: late-afternoon street corner,
red brick buildings, pedestrian zebra crossings, storefronts along the road,
soft directional afternoon sunlight, transparent natural color tones.Official Seedance 2.5 demo — ByteDance / Volcano Engine practice guide.
Again: @image1 is followed immediately by a colon and a full declaration of what it controls — the environment, down to the light direction.
Keyframe referencing is the lightest use of the syntax — two images, two declarations:
Cinematic brand concept film. @Image 1 is the first frame: the frame sways
slightly, the camera gradually pushes in toward tree shadows rushing past
the window, faster and faster — hard cut to @Image 2, the speed suddenly
eases, and the camera drifts slowly forward along a stream amid birdsong
and blooming flowers.Official Seedance 2.5 demo — ByteDance / Volcano Engine practice guide (prompt rendered in English from the guide's original).
Most video models condition on references implicitly: you attach one image and the model decides what it means. Seedance 2.5 makes the binding explicit and scalable:
| Approach | Binding | Scale | Failure mode |
|---|---|---|---|
| Plain text prompting | None — description only | 0 references | Subject drift every generation |
| Implicit image conditioning | Model guesses the image's role | Usually 1-few images | Wrong role assignment |
| Seedance 2.5 @ syntax | You declare each file's role in-line | Up to 50 files, 3 modalities | Only fails if you skip declarations |
That last cell is the honest caveat: the syntax shifts responsibility to you. Skip the usage declarations and 50 references perform worse than 5 well-declared ones.
Use the @ symbol with the material's index — @Image 1 — and immediately follow it with one sentence declaring its role, e.g. "Use the woman in @image1 as the lead character." The official guide states the model gets confused without the declaration.
Up to 50 in a single generation: 30 images, 10 video clips, and 10 audio clips. Video clips run 2-30 seconds each with a 30-second combined cap; audio has its own separate 30-second cap.
Yes — audio-only reference is new in 2.5. A music track, voice line, or sound-effect bed can drive the pacing, beat-syncing, and lip alignment of the generated video with no visual reference attached.
Almost always because references were listed without role declarations, or a multi-character cast was uploaded without an explicit who-is-who mapping. Declare every @ reference and, past five subjects, use one single-view image per character.
Yes. The model natively supports 10-plus languages including English, Chinese, Spanish, Japanese, and Arabic, and the @ syntax works identically in all of them.
A reference prompt generates new footage guided by your materials (text instructions + @ references). An editing prompt feeds an existing clip as @Video 1, describes only the local change, and the model auto-locks the source's aspect ratio and duration.
The @ reference syntax is the real interface of Seedance 2.5 — the 30-second one-shots and 4K output get the headlines, but the 50-file reference stack is what makes outputs repeatable and directable. The whole system reduces to one habit: never write an @ without a declaration, and map every character explicitly once your cast grows. Ratio your materials by the guide's priority order instead of maxing out the slots, and you get the consistency the wuxia and ensemble demos show. The fastest way to build the habit is reps: open Seedance 2.5 on Vogoo — free to try, up to 4K, native synced audio, up to 50 references per job — start with one subject image plus one declared motion reference, and scale up from there.
Note on sourcing: reference specs, the @ declaration rule, ratio recommendations, and all quoted prompts come from ByteDance's official enterprise practice guide (primary source above); launch facts are cross-checked against the Seed blog and independent media. Model API pricing for Seedance 2.5 has not been published at the time of writing, so this guide states none.


Use an AI picture generator of yourself: upload one selfie, pick a style, and generate realistic, anime, or avatar images from your own photo. Free to try.


Learn how to extend AI videos with Seedance 2.5: chained continuation prompts, official ByteDance case studies, and when a 30s one-shot beats extension.


How to find the aspect ratio of your screen or image: 3 ways — system settings, file info and editor settings — plus a resolution-to-ratio table.

Newsletter
Subscribe to our newsletter for the latest news and updates