30 秒影片,50+ 參考圖。 立即體驗
Seedance 2.5 使用指南

Seedance 2.5 使用指南

使用 Seedance 2.5 創作 AI 影片的完整指南。

Seedance 2.5 是 ByteDance Seed 新一代 AI 影片模型,具備 30 秒影片生成、50 個參考輸入、更佳一致性與進階編輯功能。

總覽

什麼是 Seedance 2.5?

Seedance 2.5 是由 ByteDance Seed 開發的新一代 AI 影片生成模型,於 2026 年 6 月 23 日發布,並於 2026 年 8 月 7 日在 SeeGen AI 上線。

作為 Seedance 2.0 的後續版本,它在影片時長、參考控制、一致性與編輯功能方面均有提升。

Seedance 2.5 功能總覽
影片時長
單則影片最長可生成 30 秒。
參考輸入
最多支援 50 個參考素材:30 張圖片+10 部影片+10 個音訊檔案。
圖片參考
最多 30 張圖片;支援最高 4K 解析度;每張圖片最大 30MB。
影片參考
最多 10 部影片;支援 480p–4K;每部影片最大 200MB;單部影片時長 2–30 秒;總時長 ≤30 秒。
音訊參考
最多 10 個音訊檔案;每個檔案最大 15MB;單個音訊時長 2–30 秒;總時長 ≤30 秒。
主要提升
更長的影片生成時長、更強的參考控制能力、更佳的一致性、進階影片編輯功能,以及獨立的音訊參考輸入。
主要功能

探索 Seedance 2.5 的主要功能

Seedance 2.5 帶來更長的影片生成、更豐富的多模態參考、精準的影片編輯、獨立音訊參考,以及更廣泛的語言支援。

01參考素材

多重參考生成

單次生成最多可使用 50 個參考素材:30 張圖片、10 部影片與 10 個音訊檔案。結合人物圖片、產品照片、動作參考、場景參考與音訊,讓複雜影片中的主體、動作、環境與視覺風格更加一致。

02時長

最長 30 秒影片生成

單次生成最長可達 30 秒的影片,是 Seedance 2.0 上限 15 秒的兩倍。更長的時長讓您有更多空間呈現多步驟動作、產品展示、對話、鏡頭運動與完整故事情節,無需拆分成多個片段。

03編輯

更精準的影片編輯控制

編輯現有影片的特定部分,而不必重新生成整個場景。Seedance 2.5 可替換物件、更改視覺細節、調整角色或修改場景的部分內容,同時保留原始構圖、光線、動態、音訊與時間軸連貫性等元素。

04音訊

音訊驅動創作

將音訊作為獨立參考使用,不需要圖片或影片參考。您最多可上傳 10 個音訊檔案,每個檔案時長 2–30 秒,合計音訊長度最長 30 秒,用以引導節奏、對話、氛圍、步調與音畫同步時機。

05語言支援

支援 10+ 種語言

以 10 種以上語言製作影片,包括中文、英文、西班牙文、印尼文、馬來文、泰文、阿拉伯文、葡萄牙文、越南文、日文和韓文。這讓 Seedance 2.5 更適合用於在地化廣告、多語言敘事、產品影片,以及面向國際受眾製作的內容。

06影片延伸

影片延伸

向前或向後延伸現有影片,讓場景延續超出原始片段。Seedance 2.5 的設計著重保留主體、光線、動態與視覺風格等關鍵元素,讓您更容易利用手邊素材製作更長的影片序列。

提示詞指南

如何寫出好的 Seedance 2.5 提示詞

好的 Seedance 2.5 提示詞應清楚描述您想呈現的畫面、主體該如何運動,以及最終影片應具備的風格。

01

主體

02

動作

03

環境

04

光線

05

運鏡

06

視覺風格

07

品質

08

限制條件

主體和動作為必要項目,其餘元素可依需求選擇加入。若提示詞較長,建議加入簡單的時間軸,讓動作與運鏡更清楚。

提示詞範例
Create a 30-second ultra-realistic candid home-video of a young Korean woman spending a quiet late morning in a residential Korean neighborhood.

SUBJECT:
Young Korean woman, early 20s, natural skin, minimal makeup, black wavy hair in a messy side ponytail with wispy bangs. Faded charcoal sleeveless crop top, loose light-wash jeans, black canvas sneakers, black cord necklace. Keep her face, body, hairstyle, and clothing consistent throughout.

SETTING:
Quiet Korean residential alleys with low-rise homes, terraces, potted plants, laundry lines, bicycles, utility poles, overhead wires, and mature trees. No shops, cafés, crowds, or commercial activity.

STYLE:
Early-2000s consumer DV camcorder footage. Imperfect handheld shake, awkward framing, autofocus hunting, exposure shifts, slight motion blur, faded colors, soft contrast, compression noise, and occasional reframing. No stabilization, cinematic camera moves, or modern color grading. It should feel genuinely recorded, not AI-generated.

TIMELINE:
00:00–00:05 — She sits outside her house adjusting her ponytail, smiling casually as the camera struggles to focus.
00:05–00:10 — She walks into a narrow alley, crouches to pet and feeds a stray cat. Focus shifts between her and the cat.
00:10–00:15 — She hangs laundry in a small front yard. Fabric moves in the breeze as exposure changes naturally.
00:15–00:20 — She sits on a terrace holding a ceramic coffee cup, watching the neighborhood and brushing hair behind her ear.
00:20–00:25 — Side profile. Someone off-camera greets her. She turns, smiles, waves, and naturally says, "Annyeong."
00:25–00:30 — She walks down a tree-lined lane with her coffee, notices the camera, gives a small smile, looks away, and keeps walking. Abrupt cut to black mid-motion.

AUDIO:
Natural location sound only: birds, distant motorcycles, wind, leaves, footsteps, faint neighborhood chatter, cat sounds, and laundry movement. Natural Korean speech only. No music, narration, or cinematic sound effects.

GOAL:
Make it feel like a forgotten early-2000s personal home video: intimate, spontaneous, imperfect, warm, and believable. Prioritize realistic motion, natural expressions, environmental detail, and identity consistency.
提示詞範例成果
應用案例

看看您能用 Seedance 2.5 創作什麼

探索真實的 Seedance 2.5 範例、可直接使用的提示詞,以及適用於不同影片工作流程的創意應用案例。

電影感

Create a 20-second, 16:9 photoreal live-action sci-fi sequence with SEVEN slow, deliberate shots. STYLE: Large-format anamorphic cinema. Heavy fog, slow falling snow, pale ice-blue, bone-white, and wet slate-grey palette with almost no saturation. The ONLY accent color is molten gold. Flat overcast light, no sun or hard shadows. Fine film grain, restrained grade, high-budget sci-fi epic. Original character and creature design only. CHARACTER LOCK: One slight lone figure in a floor-length bone-white felt coat with a deep fur-lined hood. Frost on shoulders and hem, ash-white wet hair, pale skin, frost on eyelashes, dark inner layers, worn gloves. No armor, weapons, insignia, jewelry, piercings, or forehead markings. Eyes stay closed until shot 4, then glow molten gold. COLOSSUS LOCK: An immense kneeling stone figure half-buried in ice. Only the bowed head, one shoulder, and one open palm are visible. Weathered granite, erosion, moss, and ice. Smooth ruined face, eyes closed, no horns, teeth, or demonic features. Thin dormant gold veins run through the cracks. WORLD: A frozen drowned city with ruined towers, snapped bridges, industrial pipes, and cables fading into white fog. No people, text, or signage. SHOTS: 0–3s — Extreme wide: frozen city, giant bowed head and open palm rising from the ice. Snow falls steadily. 3–6s — Wide slow push-in: the tiny figure stands on the colossus's palm, smaller than one fingernail. 6–9s — Close portrait: eyes closed, frost on lashes, visible breath, soft white fog behind. 9–12s — Extreme close-up: eyes open and glow deep molten gold, the first and only strong color. 12–15s — Wide low angle: the figure stands on a hovering dark stone slab above the city, coat blowing in the wind, colossus behind. 15–17.5s — Medium: the figure slowly raises one gloved hand toward the city below. 17.5–20s — Extreme wide: molten-gold veins spread through the ice and colossus as towers, bridges, and cables silently rise into the fog. CUTS: Hard cuts at 3s, 6s, 9s, 12s, 15s, and 17.5s. No fades, dissolves, speed ramps, slow motion, or opening freeze. Continuous subtle motion in every shot: snow, fog, breath, and fabric. AUDIO: Low wind, groaning ice, snow on fabric, sub-bass swell when the eyes open, rising low hum as the gold spreads, deep grinding as the city rises. No voice, narration, dialogue, or music. NEGATIVE: Anime, manga, stylized animation, ninja imagery, patterned cloaks, ringed irises, orange/spiky hair, weapons, gore, crowds, warm sunlight, blue sky, saturated colors except gold, lens flares, text, logos, watermark, cartoon, 3D render, video-game cinematic look, slow motion, speed ramping.

短片

@Image1 as the only protagonist. Strictly preserve the protagonist's face, hairstyle, skin tone, age, and visual identity throughout. Create a 30-second cinematic emotional confrontation. This is not an action scene or gunfight. The entire scene focuses on the protagonist's painful internal struggle before pulling the trigger. SETTING: A quiet, dark, oppressive interior such as an abandoned room, hallway, warehouse, or nighttime space. Keep the background simple and unobtrusive. Moody cinematic lighting with clear facial contrast so the protagonist's eyes, tears, and subtle expressions remain visible. CAMERA: Over-the-shoulder shot from behind the other person. Never show their face; only a blurred shoulder, back of the head, or silhouette in the foreground. The camera faces the protagonist directly, with the gun aimed toward the camera. Stable cinematic handheld with subtle breathing and a very slow push-in. Keep focus on the protagonist's face, eyes, trembling hand, and gun. PERFORMANCE: The protagonist keeps the gun raised throughout. He is angry, hurt, resentful, heartbroken, and forcing himself to act, but clearly cannot bring himself to shoot. His eyes are red and filled with tears, breathing grows heavier, lips tremble slightly, jaw stays tense, and his hand subtly shakes. Several times his finger tightens as if he is about to pull the trigger, but he stops at the last moment. His expression should constantly shift between anger, pain, attachment, resentment, sorrow, and reluctance. He is not cold-blooded and not simply afraid — the feeling is: "I have to do this, but I cannot make myself do it." Keep the acting restrained and realistic. No screaming, dramatic crying, or exaggerated movement. Tears gather but do not fully fall apart. The emotional power comes from someone barely holding themselves together. ENDING: End unresolved. He is still aiming the gun, still unable to pull the trigger, breathing unevenly with tear-filled eyes. Hold on this suspended high-tension moment. STYLE: Cinematic emotional close-up, restrained performance, subtle facial acting, trembling hands, tear-filled eyes, high-tension silence, dramatic over-the-shoulder composition. Minimal cuts and minimal physical movement. Do not turn it into an action movie. AUDIO: Real environmental sound only: faint room tone, distant reverb, uneven breathing, soft clothing movement, and subtle hand/body movement. No music, dialogue, narration, subtitles, text, or UI.

產品影片

Create a 30-second tutorial video showing how to set up and use a capsule coffee machine. 0–2s — Title card: "Seedance Capsule Coffee Machine Setup Tutorial." 2–5s — Step 1: Install the water tank. Use @image1. Slightly high rear angle. Align the tank with the back slot and push down until it clicks. VO: "First, install the water tank. Align it with the back slot and push until it clicks." 5–9s — Step 2: Install the drip tray. Use @image2. Front close-up. Slide the tray into the bottom rails until fully seated. VO: "Next, install the drip tray by sliding it into the bottom rails." 9–13s — Step 3: Install the used-capsule box. Use @image3. Close-up from a low angle. Push the box into the cavity beneath the drip tray. VO: "Then insert the capsule collection box. Used capsules will drop here automatically." 13–18s — Step 4: Fill the water tank. Use @image4. Side close-up. Open the lid, fill with clean water to the MAX line, then close it. VO: "Fill the tank with clean water, but do not exceed the MAX line." 18–25s — Step 5: Power on. Use @image5. Front medium shot. Plug in the machine and press the power button. Show the indicator blinking, then turning steady. VO: "Press the power button. When the blinking light turns steady, the machine is ready." 25–30s — Step 6: First rinse. Use @image6. Front-side view. Without inserting a capsule, press the brew button and let hot water rinse the system. VO: "For the first rinse, do not insert a capsule. Press brew, and once rinsing is complete, the machine is ready to use."

商業廣告

Create a bright, colorful commercial for fruity cookies in four flavors: strawberry, apple, grape, and orange. Use @image1 for the strawberry flavor reference. STYLE: Clean, premium, high-energy product ad with geometric cookie arrangements, vivid fruit colors, fast rhythmic editing, and strong musical timing. SEQUENCE: - Open with fruits rapidly orbiting around the hero cookie, inspired by @video1. - Different cookie flavors spiral toward the camera, switching on the beat, inspired by @video2. - Cookie arrays pan left and right with fast flavor changes, inspired by @video3. - Add vertical up-and-down movement like a precise machine, inspired by @video4. - Climax: a cookie snaps in half in slow motion, filling bursts out and crumbs scatter, inspired by @video5. - End with all four flavors lined up neatly, fruits bouncing in sync, and the text "Fresh on Seedance, made for viral vision" appearing word by word, inspired by @video6. Keep the visuals colorful, delicious, youthful, energetic, and highly shareable.

UGC 影片

Create a 30-second single-take handheld boyfriend travel vlog in Tokyo. Use the woman from @Image1 as the main character. Keep her face, hairstyle, body proportions, and overall appearance fully consistent throughout. STYLE: Authentic personal phone or small-camera vlog footage. Natural handheld shake, imperfect framing, autofocus shifts, slight motion blur, and realistic exposure changes. Casual, spontaneous, and non-cinematic. She does not pose and sometimes forgets the camera is there. TIMELINE: 0–5s — Morning in a small Tokyo apartment. She fixes her hair near the bed, notices the camera, smiles, laughs, and playfully tells her boyfriend to stop filming. 5–10s — They walk through a quiet Tokyo neighborhood. She enters a convenience store, looks at drinks and snacks, then turns and asks which one she should choose. 10–18s — At a small local restaurant, she eats ramen with chopsticks, reacts naturally to the hot food, and laughs while the boyfriend laughs behind the camera. 18–25s — They explore a lively Tokyo neighborhood. She browses shops, takes photos, and occasionally looks back at the camera while crowds pass naturally. 25–30s — Tokyo at night. She walks through illuminated streets, turns back and smiles, then sits by a train window watching city lights pass outside. REQUIREMENTS: Natural expressions, realistic movement, authentic Tokyo ambience, stable identity, and continuous handheld vlog feeling. No commercial style, dramatic posing, text, logos, CGI look, artificial transitions, or identity changes.

音樂影片

Cinematic hip-hop / rap music video, photoreal quality, high-end tone, seaside setting. Build the frame from @image1: a band performs at a golden sand beach with crashing waves — a lead vocalist gripping a mic on a stand in the wet sand, one guitarist left, one right, a drummer at the back; a vast coastline behind, rolling waves, a warm golden-hour sun shimmering on the water, sea mist in the air. The lead in a red tracksuit raps to camera — lips and jaw precisely synced to every word, head punching to the beat. Bright, punchy, fast, confident rap. HARD CUT on the beat, each switch a double contrast (shot size and type change together). Lyrics (the lead sings 'hello' in each language in turn, precisely lip-synced): English "Hello", Chinese "你好", Japanese "こんにちは", Korean "안녕하세요", Portuguese "Olá", Thai "สวัสดี", Spanish "Hola", Arabic "مرحبا". 8 hard-cut shots (low-angle wide establishing; close-up rap to camera; macro guitar-string insert; 3/4 prowling orbit; lateral track at shore; drummer tilt-up; tight push on the lead; heroic full-band push-in), one language per shot. White balance 4000K, teal-and-amber grade, 35mm, shallow depth of field, film grain, sea mist, golden-hour flare. Premium feel, precise lip-sync, no subtitles, no text overlays, hard cuts only, total 20 seconds.

模型比較

Seedance 2.5 對比 Seedance 2.0

Seedance 2.5 提供更長的影片時長、更多參考素材與更強的編輯能力,而 Seedance 2.0 則仍是短影片中速度更快、價格更實惠的選擇。

功能Seedance 2.5Seedance 2.0
狀態已在 SeeGen AI 上線已在 SeeGen AI 提供
主要重點更長的影片與進階創意控制品質、控制與成本的平衡
最適合短劇、廣告、音樂影片與複雜敘事日常 AI 影片、廣告與創意專案
影片品質最佳整體品質與場景連貫性穩定可靠的影片品質
角色一致性更強的角色、臉部與場景一致性多數標準場景下有良好一致性
動作控制進階動作、鏡頭與參考導向控制自然動作,具備扎實的提示詞控制
影片長度最長 30 秒最長 15 秒
多模態參考最多 30 張圖片、10 部影片與 10 個音訊參考最多 9 張圖片、3 部影片與 3 個音訊參考
獨立音訊輸入支援不支援
影片編輯精準影片編輯與參考轉影片工作流程支援影片編輯、延伸與參考工作流程
價格進階功能對應高階定價標準定價
API 存取SeeGen AI 已提供SeeGen AI 已提供
現在就需要的最佳選擇適合需要最大控制力、更長影片與複雜專案適合追求品質、功能與價格平衡
使用方法

如何在 SeeGen AI 上使用 Seedance 2.5

1

上傳素材

生成前請上傳您的參考圖片、影片或音訊。真人素材可能需要審核,公眾人物、受版權/智慧財產保護的角色或受限內容則可能被拒絕。

2

輸入提示詞並生成

選擇您的參考素材,撰寫描述您想要的影片的提示詞,選擇時長與其他設定,然後點擊「生成」提交您的任務。

3

分享或下載

預覽完成的影片並下載到您的裝置、複製分享連結,或在 SeeGen AI Discord 社群與其他社群平台分享您的創作。

常見問題

關於 Seedance 2.5 AI 影片生成器的常見問題

我可以在 SeeGen AI 上免費試用 Seedance 2.5 嗎?+
可以。註冊並加入我們的 Discord 社群即可獲得 200 點免費點數,足以試用一次 480p、6 秒的 Seedance 2.5 影片。
Seedance 2.5 支援真人參考素材嗎?+
支援。SeeGen AI 允許使用真人圖片與影片作為 Seedance 2.5 的參考素材,但這些素材須先通過審核才能使用。
為什麼我的素材通過審核後,生成仍然失敗?+
素材審核與影片生成分屬不同的審核階段。若生成內容被偵測到含有敏感、受限或受版權保護的內容,任務仍可能被攔截。任務失敗時,點數會自動退還。
為什麼我無法在 Seedance 2.5 多參考模式中編輯影片?+
與 Seedance 2.0 不同,Seedance 2.5 可能會將包含編輯指令的提示詞判定為影片編輯任務,導致多參考生成失敗。若要編輯現有影片,建議改用專屬的影片編輯模式。
我可以只用音訊而不搭配圖片或影片參考嗎?+
可以。Seedance 2.5 支援將音訊作為獨立參考素材。您可以用音樂、人聲或其他音訊來引導節奏、時序、氛圍與音畫同步。
Seedance 2.5 最多可以使用多少個參考素材?+
Seedance 2.5 單次生成最多支援 50 個參考素材:30 張圖片、10 部影片、10 個音訊檔。參考影片總時長最長可達 30 秒,音訊則另計。
影片最長可以生成多久?+
Seedance 2.5 最長可生成 30 秒的影片,相較之下 Seedance 2.0 上限為 15 秒。
Seedance 2.5 在 SeeGen AI 上支援哪些解析度?+
Seedance 2.5 原生支援 480p、720p 與 1080p 生成。在 SeeGen AI 上,您也可以透過自動放大功能取得 2K 或 4K 輸出。
即刻創作

在 SeeGen AI 上用 Seedance 2.5 開始創作。

透過多參考輸入、音訊引導與進階編輯,把您的創意化為更長、更具一致性的 AI 影片,一站搞定!