30秒動画、50以上の参照素材に対応。 今すぐ試す
Seedance 2.5 ユーザーガイド

Seedance 2.5 ユーザーガイド

Seedance 2.5 でAI動画を制作するための完全ガイド。

Seedance 2.5 は ByteDance Seed の次世代AI動画モデルです。最長30秒の動画生成、最大50件の参照入力、向上した一貫性、高度な編集機能を備えています。

概要

Seedance 2.5 とは?

Seedance 2.5 は ByteDance Seed が開発した次世代のAI動画生成モデルです。2026年6月23日に発表され、2026年8月7日に SeeGen AI で提供を開始しました。

Seedance 2.0 の後継として、動画の長さ、参照コントロール、一貫性、編集機能のすべてが進化しています。

Seedance 2.5 機能一覧
動画の長さ
1本あたり最長30秒の動画を生成できます。
参照入力
最大50件の参照に対応:画像30枚 + 動画10本 + 音声10ファイル。
参照画像
最大30枚。最大4K解像度に対応。1枚あたり最大30MB。
参照動画
最大10本。480p〜4Kに対応。1本あたり最大200MB。各動画2〜30秒、合計時間は30秒以内。
参照音声
最大10ファイル。1ファイルあたり最大15MB。各音声2〜30秒、合計時間は30秒以内。
主な改善点
より長い動画生成、より強力な参照コントロール、向上した一貫性、高度な動画編集機能、そして独立した音声参照入力。
主な機能

Seedance 2.5 の主な機能

Seedance 2.5 は、より長い動画生成、より豊富なマルチモーダル参照、精密な動画編集、単独の音声参照、より幅広い言語サポートを実現します。

01参照

マルチリファレンス生成

1回の生成で最大50件の参照素材(画像30枚、動画10本、音声10ファイル)を使用できます。キャラクター画像、商品写真、モーション参照、シーン参照、音声を組み合わせることで、複雑な動画でも被写体・動作・環境・ビジュアルスタイルの一貫性を高められます。

02長さ

最長30秒の動画生成

1回の生成で最長30秒の動画を作成できます。これは Seedance 2.0 の上限15秒の2倍です。長尺化により、複数ステップの動作、商品デモ、会話、カメラワーク、完結したストーリーを、複数クリップに分割せずに表現できます。

03編集

より自在な動画編集

シーン全体を作り直すことなく、既存動画の特定の部分だけを編集できます。Seedance 2.5 は、元の構図、ライティング、動き、音声、タイムラインの連続性を保ちながら、オブジェクトの置き換え、ビジュアルディテールの変更、キャラクターの調整、シーンの一部の修正が可能です。

04音声

音声主導のクリエイション

画像や動画の参照なしで、音声を単独の参照として使用できます。最大10ファイル(各2〜30秒、合計30秒以内)の音声をアップロードし、リズム、セリフ、ムード、テンポ、映像と音のタイミングをコントロールできます。

05言語

10以上の言語に対応

中国語、英語、スペイン語、インドネシア語、マレー語、タイ語、アラビア語、ポルトガル語、ベトナム語、日本語、韓国語を含む10以上の言語で動画を作成できます。これにより、Seedance 2.5はローカライズ広告、多言語ストーリーテリング、商品動画、海外向けコンテンツ制作により実用的になります。

06延長

動画の延長

既存の動画を前後に延長し、元のクリップの先までシーンを続けられます。Seedance 2.5 は被写体、ライティング、動き、ビジュアルスタイルなどの重要な要素を保持するよう設計されており、手持ちの映像から長尺シーケンスを簡単に作れます。

プロンプトガイド

良い Seedance 2.5 プロンプトの書き方

良い Seedance 2.5 プロンプトは、何を見せたいか、被写体がどう動くべきか、完成動画がどんなスタイルであるべきかを明確に記述します。

01

被写体

02

動作

03

環境

04

ライティング

05

カメラワーク

06

ビジュアルスタイル

07

品質

08

制約条件

被写体と動作は必須で、その他の要素は必要に応じて追加します。長いプロンプトでは、簡単なタイムラインを加えると動作やカメラワークがより明確になります。

プロンプト例
Create a 30-second ultra-realistic candid home-video of a young Korean woman spending a quiet late morning in a residential Korean neighborhood.

SUBJECT:
Young Korean woman, early 20s, natural skin, minimal makeup, black wavy hair in a messy side ponytail with wispy bangs. Faded charcoal sleeveless crop top, loose light-wash jeans, black canvas sneakers, black cord necklace. Keep her face, body, hairstyle, and clothing consistent throughout.

SETTING:
Quiet Korean residential alleys with low-rise homes, terraces, potted plants, laundry lines, bicycles, utility poles, overhead wires, and mature trees. No shops, cafés, crowds, or commercial activity.

STYLE:
Early-2000s consumer DV camcorder footage. Imperfect handheld shake, awkward framing, autofocus hunting, exposure shifts, slight motion blur, faded colors, soft contrast, compression noise, and occasional reframing. No stabilization, cinematic camera moves, or modern color grading. It should feel genuinely recorded, not AI-generated.

TIMELINE:
00:00–00:05 — She sits outside her house adjusting her ponytail, smiling casually as the camera struggles to focus.
00:05–00:10 — She walks into a narrow alley, crouches to pet and feeds a stray cat. Focus shifts between her and the cat.
00:10–00:15 — She hangs laundry in a small front yard. Fabric moves in the breeze as exposure changes naturally.
00:15–00:20 — She sits on a terrace holding a ceramic coffee cup, watching the neighborhood and brushing hair behind her ear.
00:20–00:25 — Side profile. Someone off-camera greets her. She turns, smiles, waves, and naturally says, "Annyeong."
00:25–00:30 — She walks down a tree-lined lane with her coffee, notices the camera, gives a small smile, looks away, and keeps walking. Abrupt cut to black mid-motion.

AUDIO:
Natural location sound only: birds, distant motorcycles, wind, leaves, footsteps, faint neighborhood chatter, cat sounds, and laundry movement. Natural Korean speech only. No music, narration, or cinematic sound effects.

GOAL:
Make it feel like a forgotten early-2000s personal home video: intimate, spontaneous, imperfect, warm, and believable. Prioritize realistic motion, natural expressions, environmental detail, and identity consistency.
プロンプト例の生成結果
活用例

Seedance 2.5 で作れるもの

Seedance 2.5 の実例、すぐに使えるプロンプト、さまざまな動画ワークフロー向けの活用例をご覧ください。

シネマティック

Create a 20-second, 16:9 photoreal live-action sci-fi sequence with SEVEN slow, deliberate shots. STYLE: Large-format anamorphic cinema. Heavy fog, slow falling snow, pale ice-blue, bone-white, and wet slate-grey palette with almost no saturation. The ONLY accent color is molten gold. Flat overcast light, no sun or hard shadows. Fine film grain, restrained grade, high-budget sci-fi epic. Original character and creature design only. CHARACTER LOCK: One slight lone figure in a floor-length bone-white felt coat with a deep fur-lined hood. Frost on shoulders and hem, ash-white wet hair, pale skin, frost on eyelashes, dark inner layers, worn gloves. No armor, weapons, insignia, jewelry, piercings, or forehead markings. Eyes stay closed until shot 4, then glow molten gold. COLOSSUS LOCK: An immense kneeling stone figure half-buried in ice. Only the bowed head, one shoulder, and one open palm are visible. Weathered granite, erosion, moss, and ice. Smooth ruined face, eyes closed, no horns, teeth, or demonic features. Thin dormant gold veins run through the cracks. WORLD: A frozen drowned city with ruined towers, snapped bridges, industrial pipes, and cables fading into white fog. No people, text, or signage. SHOTS: 0–3s — Extreme wide: frozen city, giant bowed head and open palm rising from the ice. Snow falls steadily. 3–6s — Wide slow push-in: the tiny figure stands on the colossus's palm, smaller than one fingernail. 6–9s — Close portrait: eyes closed, frost on lashes, visible breath, soft white fog behind. 9–12s — Extreme close-up: eyes open and glow deep molten gold, the first and only strong color. 12–15s — Wide low angle: the figure stands on a hovering dark stone slab above the city, coat blowing in the wind, colossus behind. 15–17.5s — Medium: the figure slowly raises one gloved hand toward the city below. 17.5–20s — Extreme wide: molten-gold veins spread through the ice and colossus as towers, bridges, and cables silently rise into the fog. CUTS: Hard cuts at 3s, 6s, 9s, 12s, 15s, and 17.5s. No fades, dissolves, speed ramps, slow motion, or opening freeze. Continuous subtle motion in every shot: snow, fog, breath, and fabric. AUDIO: Low wind, groaning ice, snow on fabric, sub-bass swell when the eyes open, rising low hum as the gold spreads, deep grinding as the city rises. No voice, narration, dialogue, or music. NEGATIVE: Anime, manga, stylized animation, ninja imagery, patterned cloaks, ringed irises, orange/spiky hair, weapons, gore, crowds, warm sunlight, blue sky, saturated colors except gold, lens flares, text, logos, watermark, cartoon, 3D render, video-game cinematic look, slow motion, speed ramping.

ショートフィルム

@Image1 as the only protagonist. Strictly preserve the protagonist's face, hairstyle, skin tone, age, and visual identity throughout. Create a 30-second cinematic emotional confrontation. This is not an action scene or gunfight. The entire scene focuses on the protagonist's painful internal struggle before pulling the trigger. SETTING: A quiet, dark, oppressive interior such as an abandoned room, hallway, warehouse, or nighttime space. Keep the background simple and unobtrusive. Moody cinematic lighting with clear facial contrast so the protagonist's eyes, tears, and subtle expressions remain visible. CAMERA: Over-the-shoulder shot from behind the other person. Never show their face; only a blurred shoulder, back of the head, or silhouette in the foreground. The camera faces the protagonist directly, with the gun aimed toward the camera. Stable cinematic handheld with subtle breathing and a very slow push-in. Keep focus on the protagonist's face, eyes, trembling hand, and gun. PERFORMANCE: The protagonist keeps the gun raised throughout. He is angry, hurt, resentful, heartbroken, and forcing himself to act, but clearly cannot bring himself to shoot. His eyes are red and filled with tears, breathing grows heavier, lips tremble slightly, jaw stays tense, and his hand subtly shakes. Several times his finger tightens as if he is about to pull the trigger, but he stops at the last moment. His expression should constantly shift between anger, pain, attachment, resentment, sorrow, and reluctance. He is not cold-blooded and not simply afraid — the feeling is: "I have to do this, but I cannot make myself do it." Keep the acting restrained and realistic. No screaming, dramatic crying, or exaggerated movement. Tears gather but do not fully fall apart. The emotional power comes from someone barely holding themselves together. ENDING: End unresolved. He is still aiming the gun, still unable to pull the trigger, breathing unevenly with tear-filled eyes. Hold on this suspended high-tension moment. STYLE: Cinematic emotional close-up, restrained performance, subtle facial acting, trembling hands, tear-filled eyes, high-tension silence, dramatic over-the-shoulder composition. Minimal cuts and minimal physical movement. Do not turn it into an action movie. AUDIO: Real environmental sound only: faint room tone, distant reverb, uneven breathing, soft clothing movement, and subtle hand/body movement. No music, dialogue, narration, subtitles, text, or UI.

商品動画

Create a 30-second tutorial video showing how to set up and use a capsule coffee machine. 0–2s — Title card: "Seedance Capsule Coffee Machine Setup Tutorial." 2–5s — Step 1: Install the water tank. Use @image1. Slightly high rear angle. Align the tank with the back slot and push down until it clicks. VO: "First, install the water tank. Align it with the back slot and push until it clicks." 5–9s — Step 2: Install the drip tray. Use @image2. Front close-up. Slide the tray into the bottom rails until fully seated. VO: "Next, install the drip tray by sliding it into the bottom rails." 9–13s — Step 3: Install the used-capsule box. Use @image3. Close-up from a low angle. Push the box into the cavity beneath the drip tray. VO: "Then insert the capsule collection box. Used capsules will drop here automatically." 13–18s — Step 4: Fill the water tank. Use @image4. Side close-up. Open the lid, fill with clean water to the MAX line, then close it. VO: "Fill the tank with clean water, but do not exceed the MAX line." 18–25s — Step 5: Power on. Use @image5. Front medium shot. Plug in the machine and press the power button. Show the indicator blinking, then turning steady. VO: "Press the power button. When the blinking light turns steady, the machine is ready." 25–30s — Step 6: First rinse. Use @image6. Front-side view. Without inserting a capsule, press the brew button and let hot water rinse the system. VO: "For the first rinse, do not insert a capsule. Press brew, and once rinsing is complete, the machine is ready to use."

CM・広告

Create a bright, colorful commercial for fruity cookies in four flavors: strawberry, apple, grape, and orange. Use @image1 for the strawberry flavor reference. STYLE: Clean, premium, high-energy product ad with geometric cookie arrangements, vivid fruit colors, fast rhythmic editing, and strong musical timing. SEQUENCE: - Open with fruits rapidly orbiting around the hero cookie, inspired by @video1. - Different cookie flavors spiral toward the camera, switching on the beat, inspired by @video2. - Cookie arrays pan left and right with fast flavor changes, inspired by @video3. - Add vertical up-and-down movement like a precise machine, inspired by @video4. - Climax: a cookie snaps in half in slow motion, filling bursts out and crumbs scatter, inspired by @video5. - End with all four flavors lined up neatly, fruits bouncing in sync, and the text "Fresh on Seedance, made for viral vision" appearing word by word, inspired by @video6. Keep the visuals colorful, delicious, youthful, energetic, and highly shareable.

UGC動画

Create a 30-second single-take handheld boyfriend travel vlog in Tokyo. Use the woman from @Image1 as the main character. Keep her face, hairstyle, body proportions, and overall appearance fully consistent throughout. STYLE: Authentic personal phone or small-camera vlog footage. Natural handheld shake, imperfect framing, autofocus shifts, slight motion blur, and realistic exposure changes. Casual, spontaneous, and non-cinematic. She does not pose and sometimes forgets the camera is there. TIMELINE: 0–5s — Morning in a small Tokyo apartment. She fixes her hair near the bed, notices the camera, smiles, laughs, and playfully tells her boyfriend to stop filming. 5–10s — They walk through a quiet Tokyo neighborhood. She enters a convenience store, looks at drinks and snacks, then turns and asks which one she should choose. 10–18s — At a small local restaurant, she eats ramen with chopsticks, reacts naturally to the hot food, and laughs while the boyfriend laughs behind the camera. 18–25s — They explore a lively Tokyo neighborhood. She browses shops, takes photos, and occasionally looks back at the camera while crowds pass naturally. 25–30s — Tokyo at night. She walks through illuminated streets, turns back and smiles, then sits by a train window watching city lights pass outside. REQUIREMENTS: Natural expressions, realistic movement, authentic Tokyo ambience, stable identity, and continuous handheld vlog feeling. No commercial style, dramatic posing, text, logos, CGI look, artificial transitions, or identity changes.

ミュージックビデオ

Cinematic hip-hop / rap music video, photoreal quality, high-end tone, seaside setting. Build the frame from @image1: a band performs at a golden sand beach with crashing waves — a lead vocalist gripping a mic on a stand in the wet sand, one guitarist left, one right, a drummer at the back; a vast coastline behind, rolling waves, a warm golden-hour sun shimmering on the water, sea mist in the air. The lead in a red tracksuit raps to camera — lips and jaw precisely synced to every word, head punching to the beat. Bright, punchy, fast, confident rap. HARD CUT on the beat, each switch a double contrast (shot size and type change together). Lyrics (the lead sings 'hello' in each language in turn, precisely lip-synced): English "Hello", Chinese "你好", Japanese "こんにちは", Korean "안녕하세요", Portuguese "Olá", Thai "สวัสดี", Spanish "Hola", Arabic "مرحبا". 8 hard-cut shots (low-angle wide establishing; close-up rap to camera; macro guitar-string insert; 3/4 prowling orbit; lateral track at shore; drummer tilt-up; tight push on the lead; heroic full-band push-in), one language per shot. White balance 4000K, teal-and-amber grade, 35mm, shallow depth of field, film grain, sea mist, golden-hour flare. Premium feel, precise lip-sync, no subtitles, no text overlays, hard cuts only, total 20 seconds.

モデル比較

Seedance 2.5 と Seedance 2.0 の比較

Seedance 2.5 はより長い動画、より多くの参照、より強力な編集機能を提供します。一方、Seedance 2.0 は短い動画向けに、より高速で手頃な選択肢であり続けます。

項目Seedance 2.5Seedance 2.0
ステータスSeeGen AI で提供中SeeGen AI で利用可能
主な強み長尺動画と高度なクリエイティブコントロール品質・コントロール・コストのバランス
最適な用途ショートドラマ、広告、ミュージックビデオ、複雑なストーリーテリング日常的なAI動画、広告、クリエイティブプロジェクト
動画品質最高水準の総合品質とシーンの連続性高品質で安定した動画
キャラクターの一貫性キャラクター・顔・シーンの一貫性がより強力標準的なシーンの多くで良好な一貫性
モーションコントロール高度なモーション・カメラ・参照ベースのコントロール自然な動きと確かなプロンプトコントロール
動画の長さ最長30秒最長15秒
マルチモーダル参照画像30枚、動画10本、音声10ファイルまで画像9枚、動画3本、音声3ファイルまで
単独の音声入力対応非対応
動画編集精密な動画編集と参照からの動画生成ワークフロー動画編集、延長、参照ワークフローに対応
料金高度な機能に応じたプレミアム価格標準価格
API アクセスSeeGen AI で利用可能SeeGen AI で利用可能
今すぐ使うならどちら?最大限のコントロール、長尺動画、複雑なプロジェクトに品質・機能・価格のバランスを重視するなら
使い方

SeeGen AI で Seedance 2.5 を使う方法

1

素材をアップロード

生成前に参照用の画像・動画・音声をアップロードします。実在の人物を含む素材は審査が必要な場合があり、著名人、著作権・IPキャラクター、制限対象のコンテンツは拒否されることがあります。

2

プロンプトを書いて生成

参照素材を選択し、作りたい動画を説明するプロンプトを書き、長さなどの設定を選んだら、「生成」をクリックしてタスクを送信します。

3

共有またはダウンロード

完成した動画をプレビューし、端末にダウンロードしたり、共有リンクをコピーしたり、SeeGen AI の Discord コミュニティや各種SNSで作品を共有できます。

よくある質問

Seedance 2.5 AI動画ジェネレーターに関するFAQ

SeeGen AI で Seedance 2.5 を無料で試せますか?+
はい。登録して Discord コミュニティに参加すると200クレジットを無料で受け取れます。480pで6秒の Seedance 2.5 動画を試すのに十分な量です。
Seedance 2.5 は実在の人物の参照に対応していますか?+
はい。SeeGen AI では、実在の人物の画像や動画を Seedance 2.5 の参照素材として使用できます。これらの素材は使用前に審査を通過する必要があります。
素材が承認されたのに生成が失敗したのはなぜですか?+
素材の承認と動画生成では審査の段階が異なります。生成内容にセンシティブ、制限対象、または著作権侵害にあたるコンテンツが検出された場合、タスクがブロックされることがあります。タスクが失敗した場合、クレジットは自動的に返金されます。
Seedance 2.5 のマルチリファレンスモードで動画を編集できないのはなぜですか?+
Seedance 2.0 と異なり、Seedance 2.5 は編集指示を含むプロンプトを動画編集タスクとして検出することがあり、その結果マルチリファレンス生成が失敗する場合があります。既存動画の編集には、専用の「動画編集」モードの使用をおすすめします。
画像や動画の参照なしで音声だけを使えますか?+
はい。Seedance 2.5 は音声を単独の参照として扱えます。音楽、ボイス、その他の音声を使って、リズム、タイミング、ムード、映像と音の同期をコントロールできます。
Seedance 2.5 では参照をいくつまで使えますか?+
Seedance 2.5 は1回の生成で最大50件の参照素材(画像30枚、動画10本、音声10ファイル)に対応します。参照動画の合計時間は最長30秒で、音声は別枠でカウントされます。
動画の最大の長さはどれくらいですか?+
Seedance 2.5 は最長30秒の動画を生成できます。Seedance 2.0 の上限は15秒です。
SeeGen AI の Seedance 2.5 はどの解像度に対応していますか?+
Seedance 2.5 は 480p、720p、1080p でネイティブ生成します。SeeGen AI では、自動アップスケールにより 2K や 4K の出力も選択できます。
今すぐ作ろう

SeeGen AI で Seedance 2.5 を使って制作しよう。

マルチリファレンス入力、音声ガイド、高度な編集で、アイデアをより長く、より一貫性のあるAI動画に。すべてがひとつの場所で!