Seedance 2.5 Prompt Guide: Multimodal Video Prompts
Last updated:
Seedance 2.5 represents a fundamental shift in AI video generation. Rather than relying solely on text prompts or single reference images, this model accepts images, videos, audio, and text as inputs—allowing you to direct every aspect of your creation like a true filmmaker. If you are new to Seedance, you can start from the Seedance 2.5.

Quick Specs
| Parameter | Specification |
|---|---|
| Image inputs | jpeg, png, webp, bmp, tiff, gif |
| Video inputs | mp4, mov reference clips |
| Audio inputs | mp3, wav reference audio |
| Text input | Natural language prompts |
| Output duration | 4–30 seconds (user-selectable) |
| Audio output | Native sound effects and music |
| Total reference limit | Up to 50 multimodal references per generation |
When working with multiple files, prioritize the assets that have the greatest impact on your final output—whether that's a reference video for motion or an image for character consistency.
How to Use @Image, @Video, and @Audio References
Seedance 2.5 uses an @ mention system to specify how each uploaded asset should be used. This gives you explicit control over what each file contributes to the generation.
Entry Points
- First/Last Frame Mode: Use when you only need a starting image plus a prompt
- Universal Reference Mode: Use for multimodal combinations (images + videos + audio + text)
Examples of Reference Instructions
| Use Case | Prompt Pattern |
|---|---|
| Set first frame | @Image1 as the first frame |
| Reference motion | Reference @Video1 for the fighting choreography |
| Copy camera work | Follow @Video1's camera movements and transitions |
| Add music/rhythm | Use @Audio1 for the background music |
| Extend a video | Extend @Video1 by 5 seconds |
| Replace character | Replace the woman in @Video1 with @Image1 |
Seedance 2.5 Prompt Formula
The @ Syntax
After uploading files, reference them in your prompt using @ followed by the file identifier:
@Image1 as the first frame, reference @Video1 for camera movement,
use @Audio1 for background music
Copy-Paste Formula
Combine subject, scene, camera, lighting, duration, and explicit @ tags in one prompt:
[Subject + action] in [setting]. [Camera move], [lighting/mood]. 30 seconds, 16:9.
@Image1 for character look · @Video1 for camera rhythm · @Audio1 for music sync
Prompt Examples for Text to Video
1Enhanced Base Quality
- Slow-motion product physics — biscuit break, filling splash, and particle impact referenced from @Video3
- Rhythmic ad editing — composition, motion, and end-card each mapped to dedicated @Video references
- Shot-by-shot instruction following — geometric flavor layout, close-ups, and tagline timing hit precise beats
- Consistent commercial look — bright palette and energetic pacing hold from opener through final brand lockup
5Creative Template Replication
- Step-by-step tutorial template — water tank, drip tray, capsule bin, fill, power, and flush each use @Image1–@Image6
- Repeatable shot language — medium, close-up, and front-angle cues scripted for every install beat
- Timed voiceover blocks — 0–2s title through 25–30s first flush with on-screen MAX line callouts
- Product-demo clarity — click-to-lock actions, indicator states, and “no capsule” flush called out explicitly
Prompt Examples for Image to Video
2Multimodal Reference System
- Lead character lock — European woman @Image1 stays recognizable through a 26-second one-shot
- Role-specific @Image slots — door, outfit, market, elephant, seasons, and fireworks each assigned separately
- Dual camera grammar — steady back-follow @Video1 blended with smooth orbital moves @Video2
- Day-to-night and season shifts inside one continuous take without identity or style drift
4Motion and Camera Replication
- Unbroken follow shot — camera tracks @Image1 protagonist left-to-right through six connected rooms
- Shared set reference @Image2 — same windows and flooring, new mood and exterior view each room
- Choreography transplant — comic fight opponent and motion borrowed from @Image3 in Room One
- Per-room camera rhythm — felt warmth, noir stop-motion, underwater drift, fireworks, then snap-to-black end card @Image8
Prompt Examples for Audio-Synced Video
3Era-Spanning Visual Storytelling
- Style references per era — ink-wash Cuju scene from @Image1, classical Greek plaza from @Image2
- One rolling ball links multimodal cuts across 3,000 years without breaking the visual thread
- Prompt-driven pacing — voiceover, era labels, and scene shifts timed to a 30-second educational arc
- Art-direction consistency — painting style changes by civilization while the ball stays the throughline
6Video Extension
- Global visual thread — one flower handed across nine countries in a fast-cut travel narrative
- Seamless locale handoffs — whip pans and out-of-frame passes link @Image1–@Image9 street scenes
- Authentic setting per reference — flower shop, market, courtyard, and urban sidewalk each from its @Image plate
- Localized performance — recipients smile and say “thank you” in native language with natural lip-sync
Prompt Examples for 3D Blockout to Video
Upload a 3D white-model blockout to lock spatial layout, then describe materials, lighting, and camera motion in text. Seedance 2.5 treats the blockout as a geometry guide while you direct the final look.
Use @Image1 as the 3D blockout layout. Animate a slow dolly through the interior — warm afternoon light through floor-to-ceiling windows, concrete floors, minimal furniture. 20 seconds, cinematic 24fps feel.
- · Keep blockout meshes clean — separate rooms and openings read best on camera
- · Name materials and lighting in the prompt; the blockout only defines structure
- · Pair with @Video1 if you need a specific camera path through the space
Common Prompt Mistakes
- Be explicit about references: Write clearly which file is for what purpose. "Reference @Video1's camera movement" is better than just mentioning the video.
- Prioritize your uploads: With a 50-reference budget, choose assets that have the greatest impact on your output.
- Check your @ mentions: With multiple files, double-check that you haven't confused which image, video, or audio goes where.
- Specify edit vs. reference: Make clear whether you want to edit an existing video or use it as a reference for generating something new.
- Duration alignment: When extending video, set your generation duration to match the new content length (e.g., extend by 5s = generate 5s).
- Use natural language: The model understands context. Describe what you want as you would to a human editor.
Seedance 2.5 Prompt Guide FAQ
Quick answers on writing Seedance 2.5 multimodal prompts — @ tags, reference limits, 30-second briefs, audio sync, 3D blockouts, and the mistakes that waste credits.
What are @ references in Seedance 2.5 prompts?
@ references are inline tags — @Image1, @Video1, @Audio1 — that tell Seedance 2.5 exactly how each uploaded file should be used. Instead of hoping the model guesses, you assign roles: @Image1 for character look, @Video1 for camera motion, @Audio1 for music rhythm. Each tag maps to one slot in your reference list, so a 30-second render can keep character, blocking, and pacing separate in one brief.
How many multimodal references can Seedance 2.5 accept?
Up to fifty mixed references on paid packs — images, video clips, audio beds, and text snippets. You do not need to fill every slot; a focused 15–25 reference brief with clear roles usually outperforms dumping fifty similar assets. Prioritize files that change look, motion, or sound most — hero character plates, one camera-reference clip, and one audio bed before decorative extras.
How do I write a Seedance 2.5 prompt for a 30-second clip?
Open with subject and action, then setting, camera move, lighting, and mood in plain language. State duration (up to 30 seconds) and aspect ratio (16:9, 9:16, 1:1, or 4:3). Close with explicit @ tags that name each reference role — for example: "@Image1 for character look · @Video1 for handheld camera rhythm · @Audio1 for beat-synced cuts." The generator shows credit cost before you submit; tighten the brief rather than stacking vague adjectives.
When should I use First/Last Frame mode vs Universal Reference mode?
Use First/Last Frame when you have a starting image (and optionally an ending frame) and mainly need text to describe motion between them — product reveals, simple image-to-video, or locked compositions. Use Universal Reference when the job mixes several asset types: multiple images, reference video for choreography or camera work, audio for rhythm, or a 3D blockout for spatial layout. Universal Reference is the default for multimodal Seedance 2.5 briefs on this guide page.
How do I copy motion or camera work from a reference video?
Upload the source clip, note its @Video label, and say what to borrow in plain language — for example: "Reference @Video1 for the fighting choreography" or "Follow @Video1's camera movements and whip-pan transitions." You can also replace subjects while keeping motion: "Replace the woman in @Video1 with @Image1." Be explicit about edit vs reference; "extend @Video1 by 5 seconds" continues the clip, while "use @Video1 as motion reference only" generates new footage with similar movement.
Can Seedance 2.5 sync video cuts to uploaded audio?
Yes. Upload a music bed or voice track, tag it @Audio1 (or @Audio2, and so on), and describe how visuals should follow the sound — beat-synced cuts, lip-sync to dialog, or ambient mood. Combine with @Video references when rhythm and camera grammar both matter. Native audio is generated with the clip; attaching @Audio tells the model which external track should drive timing and energy across the full 30-second render.
How do I prompt with a 3D white-model blockout?
Export a low-poly, untextured OBJ or FBX from Blender, Cinema 4D, Maya, or SketchUp, upload it as an image/blockout reference, and tag it in the prompt — for example: "Use @Image1 as the 3D blockout layout." Describe materials, lighting, and camera path in text; the blockout only defines geometry and framing. Pair with @Video1 if you need a specific dolly or orbit path through the space. Clean room separation and openings read best on camera.
How do I extend an existing video without breaking continuity?
Upload the clip, reference it as @Video1, and state the extension length in the prompt — for example: "Extend @Video1 by 5 seconds" with generation duration set to match the new segment. For longer chains, reuse the same @Image character and style references across segments so identity and lighting stay aligned. Review each extension before chaining again; very long chains can drift on face, wardrobe, or color grade between segments.
What are the most common Seedance 2.5 prompt mistakes?
Vague reference roles ("use the video") instead of explicit @ tags; confusing @Image3 with @Image4 in long briefs; uploading similar assets into the same role; forgetting to set duration when extending video; and describing ten scene changes without enough reference support for a 30-second single pass. Fix by naming each file's job, capping scene count to what your references can hold, and re-reading @ mentions before you hit Generate.
Can I generate from text only without uploading references?
Yes — text-to-video works with no uploads. You get the fastest brief setup, but less control over exact character identity, camera grammar, and audio timing than a multimodal job. Add @Image and @Video references when a client plate, storyboard frame, or reference clip already exists. Image-to-video and Universal Reference modes are where Seedance 2.5's 30-second, 50-reference workflow pays off for branded or narrative work.
Unleash Your Imagination With Seedance 2.5
Join the future of AI video generation with Seedance AI's powerful platform. Transform your text and images into stunning long-form cinematic videos with unprecedented ease and quality.