30s Videos, 50+ References. Try It Now
Seedance 2.5 User Guide

Seedance 2.5 User Guide

A complete guide to creating AI videos with Seedance 2.5.

Seedance 2.5 is ByteDance Seed's next-generation AI video model, featuring 30s video generation, 50 reference inputs, improved consistency, and advanced editing capabilities.

Overview

What Is Seedance 2.5?

Seedance 2.5 is a next-generation AI video generation model developed by ByteDance Seed, announced on June 23, 2026, and launched on SeeGen AI on August 7, 2026.

As the successor to Seedance 2.0, it delivers improvements in video duration, reference control, consistency, and editing capabilities.

Seedance 2.5 Feature Overview
Video Duration
Generate single videos up to 30 seconds.
Reference Inputs
Support up to 50 references: 30 images + 10 videos + 10 audio files.
Image References
Up to 30 images; supports up to 4K resolution; max 30MB per image.
Video References
Up to 10 videos; supports 480p–4K; max 200MB per video; each video 2–30 seconds; total duration ≤30 seconds.
Audio References
Up to 10 audio files; max 15MB per file; each audio 2–30 seconds; total duration ≤30 seconds.
Key Improvements
Longer video generation, stronger reference control, improved consistency, advanced video editing capabilities and independent audio reference input.
Key Features

Discover the Key Features of Seedance 2.5

Seedance 2.5 brings longer video generation, richer multimodal references, precise video editing, standalone audio references, and broader language support.

01References

Multi-Reference Generation

Use up to 50 reference assets in one generation: 30 images, 10 videos, and 10 audio files. Combine character images, product shots, motion references, scene references, and audio to keep subjects, actions, environments, and visual styles more consistent across complex videos.

02Duration

Up to 30s Video Generation

Generate videos up to 30 seconds in a single generation, twice the 15-second maximum of Seedance 2.0. The longer duration gives you more room for multi-step actions, product demonstrations, dialogue, camera movement, and complete story sequences without splitting them into multiple clips.

03Editing

Video Editing with More Control

Edit specific parts of an existing video instead of generating the whole scene again. Seedance 2.5 can replace objects, change visual details, adjust characters, or modify parts of a scene while preserving elements such as the original composition, lighting, motion, audio, and timeline continuity.

04Audio

Audio-Driven Creation

Use audio as an independent reference, without requiring an image or video reference. You can upload up to 10 audio files, with each file lasting 2–30 seconds and the combined audio length up to 30 seconds, to guide rhythm, dialogue, mood, pacing, and audiovisual timing.

05Languages

10+ Language Support

Create videos in 10+ languages, including Chinese, English, Spanish, Indonesian, Malay, Thai, Arabic, Portuguese, Vietnamese, Japanese, and Korean. This makes Seedance 2.5 more practical for localized ads, multilingual storytelling, product videos, and content made for international audiences.

06Extensions

Video Extensions

Extend an existing video forward or backward to continue the scene beyond the original clip. Seedance 2.5 is designed to preserve key elements such as the subject, lighting, motion, and visual style, making it easier to build longer sequences from footage you already have.

Prompt Guide

How to Write a Good Seedance 2.5 Prompt

A good Seedance 2.5 prompt should clearly describe what you want to see, how the subject should move, and what style the final video should have.

01

Subject

02

Action

03

Environment

04

Lighting

05

Camera Movement

06

Visual Style

07

Quality

08

Constraints

Subject and action are required, while the other elements can be added based on your needs. For longer prompts, add a simple timeline to make action and camera movement clearer.

Example Prompt
Create a 30-second ultra-realistic candid home-video of a young Korean woman spending a quiet late morning in a residential Korean neighborhood.

SUBJECT:
Young Korean woman, early 20s, natural skin, minimal makeup, black wavy hair in a messy side ponytail with wispy bangs. Faded charcoal sleeveless crop top, loose light-wash jeans, black canvas sneakers, black cord necklace. Keep her face, body, hairstyle, and clothing consistent throughout.

SETTING:
Quiet Korean residential alleys with low-rise homes, terraces, potted plants, laundry lines, bicycles, utility poles, overhead wires, and mature trees. No shops, cafés, crowds, or commercial activity.

STYLE:
Early-2000s consumer DV camcorder footage. Imperfect handheld shake, awkward framing, autofocus hunting, exposure shifts, slight motion blur, faded colors, soft contrast, compression noise, and occasional reframing. No stabilization, cinematic camera moves, or modern color grading. It should feel genuinely recorded, not AI-generated.

TIMELINE:
00:00–00:05 — She sits outside her house adjusting her ponytail, smiling casually as the camera struggles to focus.
00:05–00:10 — She walks into a narrow alley, crouches to pet and feeds a stray cat. Focus shifts between her and the cat.
00:10–00:15 — She hangs laundry in a small front yard. Fabric moves in the breeze as exposure changes naturally.
00:15–00:20 — She sits on a terrace holding a ceramic coffee cup, watching the neighborhood and brushing hair behind her ear.
00:20–00:25 — Side profile. Someone off-camera greets her. She turns, smiles, waves, and naturally says, "Annyeong."
00:25–00:30 — She walks down a tree-lined lane with her coffee, notices the camera, gives a small smile, looks away, and keeps walking. Abrupt cut to black mid-motion.

AUDIO:
Natural location sound only: birds, distant motorcycles, wind, leaves, footsteps, faint neighborhood chatter, cat sounds, and laundry movement. Natural Korean speech only. No music, narration, or cinematic sound effects.

GOAL:
Make it feel like a forgotten early-2000s personal home video: intimate, spontaneous, imperfect, warm, and believable. Prioritize realistic motion, natural expressions, environmental detail, and identity consistency.
Prompt example result
Use Cases

See What You Can Create with Seedance 2.5

Explore real Seedance 2.5 examples, ready-to-use prompts, and creative use cases for different video workflows.

Cinematic

Create a 20-second, 16:9 photoreal live-action sci-fi sequence with SEVEN slow, deliberate shots. STYLE: Large-format anamorphic cinema. Heavy fog, slow falling snow, pale ice-blue, bone-white, and wet slate-grey palette with almost no saturation. The ONLY accent color is molten gold. Flat overcast light, no sun or hard shadows. Fine film grain, restrained grade, high-budget sci-fi epic. Original character and creature design only. CHARACTER LOCK: One slight lone figure in a floor-length bone-white felt coat with a deep fur-lined hood. Frost on shoulders and hem, ash-white wet hair, pale skin, frost on eyelashes, dark inner layers, worn gloves. No armor, weapons, insignia, jewelry, piercings, or forehead markings. Eyes stay closed until shot 4, then glow molten gold. COLOSSUS LOCK: An immense kneeling stone figure half-buried in ice. Only the bowed head, one shoulder, and one open palm are visible. Weathered granite, erosion, moss, and ice. Smooth ruined face, eyes closed, no horns, teeth, or demonic features. Thin dormant gold veins run through the cracks. WORLD: A frozen drowned city with ruined towers, snapped bridges, industrial pipes, and cables fading into white fog. No people, text, or signage. SHOTS: 0–3s — Extreme wide: frozen city, giant bowed head and open palm rising from the ice. Snow falls steadily. 3–6s — Wide slow push-in: the tiny figure stands on the colossus's palm, smaller than one fingernail. 6–9s — Close portrait: eyes closed, frost on lashes, visible breath, soft white fog behind. 9–12s — Extreme close-up: eyes open and glow deep molten gold, the first and only strong color. 12–15s — Wide low angle: the figure stands on a hovering dark stone slab above the city, coat blowing in the wind, colossus behind. 15–17.5s — Medium: the figure slowly raises one gloved hand toward the city below. 17.5–20s — Extreme wide: molten-gold veins spread through the ice and colossus as towers, bridges, and cables silently rise into the fog. CUTS: Hard cuts at 3s, 6s, 9s, 12s, 15s, and 17.5s. No fades, dissolves, speed ramps, slow motion, or opening freeze. Continuous subtle motion in every shot: snow, fog, breath, and fabric. AUDIO: Low wind, groaning ice, snow on fabric, sub-bass swell when the eyes open, rising low hum as the gold spreads, deep grinding as the city rises. No voice, narration, dialogue, or music. NEGATIVE: Anime, manga, stylized animation, ninja imagery, patterned cloaks, ringed irises, orange/spiky hair, weapons, gore, crowds, warm sunlight, blue sky, saturated colors except gold, lens flares, text, logos, watermark, cartoon, 3D render, video-game cinematic look, slow motion, speed ramping.

Short Films

@Image1 as the only protagonist. Strictly preserve the protagonist's face, hairstyle, skin tone, age, and visual identity throughout. Create a 30-second cinematic emotional confrontation. This is not an action scene or gunfight. The entire scene focuses on the protagonist's painful internal struggle before pulling the trigger. SETTING: A quiet, dark, oppressive interior such as an abandoned room, hallway, warehouse, or nighttime space. Keep the background simple and unobtrusive. Moody cinematic lighting with clear facial contrast so the protagonist's eyes, tears, and subtle expressions remain visible. CAMERA: Over-the-shoulder shot from behind the other person. Never show their face; only a blurred shoulder, back of the head, or silhouette in the foreground. The camera faces the protagonist directly, with the gun aimed toward the camera. Stable cinematic handheld with subtle breathing and a very slow push-in. Keep focus on the protagonist's face, eyes, trembling hand, and gun. PERFORMANCE: The protagonist keeps the gun raised throughout. He is angry, hurt, resentful, heartbroken, and forcing himself to act, but clearly cannot bring himself to shoot. His eyes are red and filled with tears, breathing grows heavier, lips tremble slightly, jaw stays tense, and his hand subtly shakes. Several times his finger tightens as if he is about to pull the trigger, but he stops at the last moment. His expression should constantly shift between anger, pain, attachment, resentment, sorrow, and reluctance. He is not cold-blooded and not simply afraid — the feeling is: "I have to do this, but I cannot make myself do it." Keep the acting restrained and realistic. No screaming, dramatic crying, or exaggerated movement. Tears gather but do not fully fall apart. The emotional power comes from someone barely holding themselves together. ENDING: End unresolved. He is still aiming the gun, still unable to pull the trigger, breathing unevenly with tear-filled eyes. Hold on this suspended high-tension moment. STYLE: Cinematic emotional close-up, restrained performance, subtle facial acting, trembling hands, tear-filled eyes, high-tension silence, dramatic over-the-shoulder composition. Minimal cuts and minimal physical movement. Do not turn it into an action movie. AUDIO: Real environmental sound only: faint room tone, distant reverb, uneven breathing, soft clothing movement, and subtle hand/body movement. No music, dialogue, narration, subtitles, text, or UI.

Product Videos

Create a 30-second tutorial video showing how to set up and use a capsule coffee machine. 0–2s — Title card: "Seedance Capsule Coffee Machine Setup Tutorial." 2–5s — Step 1: Install the water tank. Use @image1. Slightly high rear angle. Align the tank with the back slot and push down until it clicks. VO: "First, install the water tank. Align it with the back slot and push until it clicks." 5–9s — Step 2: Install the drip tray. Use @image2. Front close-up. Slide the tray into the bottom rails until fully seated. VO: "Next, install the drip tray by sliding it into the bottom rails." 9–13s — Step 3: Install the used-capsule box. Use @image3. Close-up from a low angle. Push the box into the cavity beneath the drip tray. VO: "Then insert the capsule collection box. Used capsules will drop here automatically." 13–18s — Step 4: Fill the water tank. Use @image4. Side close-up. Open the lid, fill with clean water to the MAX line, then close it. VO: "Fill the tank with clean water, but do not exceed the MAX line." 18–25s — Step 5: Power on. Use @image5. Front medium shot. Plug in the machine and press the power button. Show the indicator blinking, then turning steady. VO: "Press the power button. When the blinking light turns steady, the machine is ready." 25–30s — Step 6: First rinse. Use @image6. Front-side view. Without inserting a capsule, press the brew button and let hot water rinse the system. VO: "For the first rinse, do not insert a capsule. Press brew, and once rinsing is complete, the machine is ready to use."

Commercial Ads

Create a bright, colorful commercial for fruity cookies in four flavors: strawberry, apple, grape, and orange. Use @image1 for the strawberry flavor reference. STYLE: Clean, premium, high-energy product ad with geometric cookie arrangements, vivid fruit colors, fast rhythmic editing, and strong musical timing. SEQUENCE: - Open with fruits rapidly orbiting around the hero cookie, inspired by @video1. - Different cookie flavors spiral toward the camera, switching on the beat, inspired by @video2. - Cookie arrays pan left and right with fast flavor changes, inspired by @video3. - Add vertical up-and-down movement like a precise machine, inspired by @video4. - Climax: a cookie snaps in half in slow motion, filling bursts out and crumbs scatter, inspired by @video5. - End with all four flavors lined up neatly, fruits bouncing in sync, and the text "Fresh on Seedance, made for viral vision" appearing word by word, inspired by @video6. Keep the visuals colorful, delicious, youthful, energetic, and highly shareable.

UGC Video

Create a 30-second single-take handheld boyfriend travel vlog in Tokyo. Use the woman from @Image1 as the main character. Keep her face, hairstyle, body proportions, and overall appearance fully consistent throughout. STYLE: Authentic personal phone or small-camera vlog footage. Natural handheld shake, imperfect framing, autofocus shifts, slight motion blur, and realistic exposure changes. Casual, spontaneous, and non-cinematic. She does not pose and sometimes forgets the camera is there. TIMELINE: 0–5s — Morning in a small Tokyo apartment. She fixes her hair near the bed, notices the camera, smiles, laughs, and playfully tells her boyfriend to stop filming. 5–10s — They walk through a quiet Tokyo neighborhood. She enters a convenience store, looks at drinks and snacks, then turns and asks which one she should choose. 10–18s — At a small local restaurant, she eats ramen with chopsticks, reacts naturally to the hot food, and laughs while the boyfriend laughs behind the camera. 18–25s — They explore a lively Tokyo neighborhood. She browses shops, takes photos, and occasionally looks back at the camera while crowds pass naturally. 25–30s — Tokyo at night. She walks through illuminated streets, turns back and smiles, then sits by a train window watching city lights pass outside. REQUIREMENTS: Natural expressions, realistic movement, authentic Tokyo ambience, stable identity, and continuous handheld vlog feeling. No commercial style, dramatic posing, text, logos, CGI look, artificial transitions, or identity changes.

Music Video

Cinematic hip-hop / rap music video, photoreal quality, high-end tone, seaside setting. Build the frame from @image1: a band performs at a golden sand beach with crashing waves — a lead vocalist gripping a mic on a stand in the wet sand, one guitarist left, one right, a drummer at the back; a vast coastline behind, rolling waves, a warm golden-hour sun shimmering on the water, sea mist in the air. The lead in a red tracksuit raps to camera — lips and jaw precisely synced to every word, head punching to the beat. Bright, punchy, fast, confident rap. HARD CUT on the beat, each switch a double contrast (shot size and type change together). Lyrics (the lead sings 'hello' in each language in turn, precisely lip-synced): English "Hello", Chinese "你好", Japanese "こんにちは", Korean "안녕하세요", Portuguese "Olá", Thai "สวัสดี", Spanish "Hola", Arabic "مرحبا". 8 hard-cut shots (low-angle wide establishing; close-up rap to camera; macro guitar-string insert; 3/4 prowling orbit; lateral track at shore; drummer tilt-up; tight push on the lead; heroic full-band push-in), one language per shot. White balance 4000K, teal-and-amber grade, 35mm, shallow depth of field, film grain, sea mist, golden-hour flare. Premium feel, precise lip-sync, no subtitles, no text overlays, hard cuts only, total 20 seconds.

Model Comparison

Seedance 2.5 vs Seedance 2.0

Seedance 2.5 offers longer videos, more references, and stronger editing, while Seedance 2.0 remains a faster and more affordable option for shorter videos.

FeatureSeedance 2.5Seedance 2.0
StatusLive on SeeGen AIAvailable on SeeGen AI
Main focusLonger videos and advanced creative controlBalanced quality, control, and cost
Best forShort dramas, ads, music videos, and complex storytellingEveryday AI videos, ads, and creative projects
Video qualityBest overall quality and scene continuityStrong, reliable video quality
Character consistencyStronger character, face, and scene consistencyGood consistency for most standard scenes
Motion controlAdvanced motion, camera, and reference-based controlNatural motion with solid prompt control
Video lengthUp to 30 secondsUp to 15 seconds
Multimodal referencesUp to 30 images, 10 videos, and 10 audio referencesUp to 9 images, 3 videos, and 3 audio references
Standalone Audio InputYesNo
Video editingPrecise video editing and reference-to-video workflowsSupports video editing, extension, and reference workflows
PricingPremium pricing for advanced featuresStandard pricing
API accessAvailable on SeeGen AIAvailable on SeeGen AI
Best choice if you need it nowChoose for maximum control, longer videos, and complex projectsChoose for a balance of quality, features, and price
How To Use

How to Use Seedance 2.5 on SeeGen AI

1

Upload Assets

Upload your reference images, videos, or audio before generation. Real-person assets may require review, while public figures, copyrighted/IP characters, or restricted content may be rejected.

2

Prompt & Generate

Select your reference assets, write a prompt describing the video you want, choose the duration and other settings, then click Generate to submit your task.

3

Share or Download

Preview the finished video and download it to your device, copy a share link, or share your creation with the SeeGen AI Discord community and other social platforms.

FAQ

FAQs about Seedance 2.5 AI Video Generator

Can I try Seedance 2.5 for free on SeeGen AI?+
Yes. Register and join our Discord community to receive 200 free credits, enough to try a 6-second Seedance 2.5 video at 480p.
Does Seedance 2.5 support real-person references?+
Yes. SeeGen AI supports real-person images and videos as reference assets for Seedance 2.5. These assets must pass review before they can be used.
Why did my generation fail after my asset was approved?+
Asset approval and video generation use different review stages. A task may still be blocked if the generation is detected as containing sensitive, restricted, or copyrighted content. If a task fails, your credits are automatically refunded.
Why can't I edit a video in Seedance 2.5 Multi-Reference mode?+
Unlike Seedance 2.0, Seedance 2.5 may detect prompts containing editing instructions as a Video Edit task, which can cause Multi-Reference generation to fail. For editing existing videos, we recommend using the dedicated Video Edit mode.
Can I use audio without an image or video reference?+
Yes. Seedance 2.5 supports audio as a standalone reference. You can use music, voice, or other audio to guide rhythm, timing, mood, and audiovisual alignment.
How many references can I use with Seedance 2.5?+
Seedance 2.5 supports up to 50 reference assets in one generation: 30 images, 10 videos, and 10 audio files. The total reference video duration can be up to 30 seconds, and audio is counted separately.
What is the maximum video length?+
Seedance 2.5 can generate videos up to 30 seconds, compared with the 15-second maximum of Seedance 2.0.
What resolutions does Seedance 2.5 support on SeeGen AI?+
Seedance 2.5 generates natively at 480p, 720p, and 1080p. On SeeGen AI, you can also request 2K or 4K output through automatic upscaling.
Create Today

Create with Seedance 2.5 on SeeGen AI.

Turn your ideas into longer, more consistent AI videos with multi-reference inputs, audio guidance, and advanced editing. All in one place!