Seedance 2.0でスタイリッシュな動画を作る鍵は、プロンプトに映画的な形容詞を増やすことではありません。人物、衣装・小物、空間・照明、開始構図を別々の参照画像に分け、それぞれから何だけを引き継ぐか明示することが重要です。
ByteDance Seedが2026年2月12日に公開した公式発表によると、Seedance 2.0はテキスト、画像、動画、音声を扱う統合マルチモーダル動画モデルです。最大9枚の画像、3本の動画、3本の音声を参照でき、公式例でも「人物は画像2、空間は画像3、小物は画像4」のように役割を指定しています。Dreaminaの現行チュートリアルでも、Multiframesモードに素材を追加し、@AssetNameで呼び出す方法が案内されています。
ただし、参照数が多いほど良いわけではありません。異なる顔、衣装、照明、構図を同時に入れると、何を固定すべきか曖昧になります。本ガイドでは一緒に使える4枚の参照画像と代表画像1枚を実際に制作し、少ない参照を明確に分担させる方法を示します。
本記事は2026年7月28日に確認したByteDance SeedとDreaminaの公式公開情報に基づきます。メニュー名、地域、プラン、クレジット、入力上限は変更される場合があります。公開ベンチマークと品質評価は開発元による結果であり、すべての素材や場面で同じ品質を保証するものではありません。
完成画より先に「参照役割表」を作る
今回の例は、銀色のショートボブ、赤いコート、半透明の傘を持つ成人パフォーマーが、雨に濡れたブルータリズム建築の通路を歩き、カメラへ向かってターンするファッションフィルムです。4枚の参照が同じ情報を重ねないよう、役割を分けます。
| 参照 | 固定する情報 | 意図的に含めない情報 |
|---|---|---|
| 画像1 | 顔、体格、髪、衣装全体 | 背景、動作、カメラ |
| 画像2 | 傘、生地、手袋、金属素材 | 新しい人物や場所 |
| 画像3 | 建築、照明、雨、反射 | 人物、衣装、小物 |
| 画像4 | 開始構図、レンズ感、色調 | 後半の動作と結末 |
一枚にすべてを入れると便利に見えますが、一部を変えたとき全体が崩れやすくなります。役割を分ければ、顔を保ったまま場所だけを変える、衣装を保ったままカメラ経路だけを変える、といった調整が容易です。
画像1:人物アンカーはポスターではなく識別情報
人物参照には、派手なポーズや顔を隠す手よりも、顔、髪、全身の輪郭、手、衣装、靴が読み取れる中立的な画像が向いています。背景を簡潔にし、モーションブラーを避けることで、固定要素と演出要素を分離できます。

画像生成プロンプト
Use case: stylized-concept
Asset type: landscape reference image for a Seedance 2.0 multi-reference fashion-film workflow; character identity anchor
Primary request: Create a premium cinematic character anchor image of one original adult female performance artist whose identity, silhouette, outfit, and materials can be reused consistently in AI video generation.
Scene/backdrop: minimal dark graphite studio with a softly illuminated floor, subtle atmospheric haze, no scenery and no distracting props.
Subject: one clearly adult woman, mid-to-late twenties, distinctive short silver bob with a blunt fringe, calm angular face, dark brown eyes, poised neutral expression. She wears a long asymmetrical crimson technical coat with a high sculpted collar over a matte black fitted bodysuit, slim black gloves, and mirror-chrome ankle boots. Full body visible, arms relaxed slightly away from the torso, outfit silhouette and footwear unobstructed.
Style/medium: photorealistic high-fashion editorial photography, original character design, realistic anatomy and natural skin texture.
Composition/framing: wide 16:9 landscape, full-body three-quarter-front view centered with generous breathing room, eye-level camera, 50mm lens character, crisp readable silhouette.
Lighting/mood: controlled soft key light from upper left, cool cyan rim light from behind, restrained crimson bounce from the coat, confident and enigmatic.
Color palette: graphite black, deep crimson, cool cyan, mirror silver, natural skin tones.
Materials/textures: matte technical fabric, subtle coat seams, brushed black textile, clean reflective chrome footwear.
Constraints: exactly one adult subject; preserve a clearly readable face, hair, coat, bodysuit, gloves, and boots; no motion blur; no text, letters, numbers, logos, trademarks, watermarks, UI, weapons, or copyrighted characters; no resemblance to a real person or living artist.
Avoid: extra people, split panels, masks, sunglasses, fantasy armor, cybernetic body parts, exaggerated anatomy, exposed underwear, cluttered background, neon cyberpunk overload, malformed hands or feet.この人物アンカーでは、銀色のショートボブ、造形的な非対称の赤いコート、マットブラックのボディスーツ、クロームブーツを固定しました。動画プロンプトでも同じ名詞を繰り返し、画像と文章の基準を合わせます。
画像2:衣装と小物を「素材サンプル」として作る
傘のように動画内で状態が変わる小物は、形、色、持ち手、素材がはっきり見える必要があります。全身画像の中で小さく写るだけでは形状が変化しやすくなります。同じ小物を複数置くと個数まで参照される可能性があるため、主役となる小物は一つだけ見せます。

画像生成プロンプト
Use case: stylized-concept
Asset type: landscape reference image for a Seedance 2.0 multi-reference fashion-film workflow; wardrobe and hero-prop detail anchor
Input images: Use the character anchor only as the identity, crimson technical coat, matte black bodysuit, glove, and chrome-boot reference. Preserve those exact design cues.
Primary request: Create a premium fashion editorial detail composition that clearly defines the performer's materials and one hero prop for video consistency.
Scene/backdrop: dark graphite studio tabletop and wall with controlled reflections, minimal and uncluttered.
Subject: the same clearly adult silver-bob performer shown in a waist-up three-quarter profile on the left, wearing the same crimson high-collar technical coat and slim black gloves; on the right, she holds exactly one closed translucent smoke-gray umbrella with a polished chrome curved handle. Include carefully composed close details of the crimson fabric seam, black glove, and chrome boot surface without split-panel borders or labels.
Style/medium: photorealistic luxury fashion campaign photography, original design, tactile material study.
Composition/framing: wide 16:9 landscape, waist-up performer occupying the left half, umbrella silhouette and material details arranged on the right, clean hierarchy and generous padding, 85mm lens character.
Lighting/mood: narrow soft key light, cool cyan edge light, restrained crimson reflection, elegant and mysterious.
Color palette: deep crimson, graphite black, smoke gray, cool cyan, mirror silver.
Materials/textures: matte technical coat fabric with fine seams, soft black glove leather, translucent umbrella canopy, polished chrome handle and boot.
Constraints: preserve the character's short silver bob, face, coat design, black bodysuit, gloves, and chrome boots from the input; show exactly one umbrella; no text, letters, numbers, labels, logos, trademarks, watermarks, UI, weapons, or copyrighted characters; no resemblance to a real person or living artist.
Avoid: duplicate props, wardrobe redesign, different hair length, open umbrella blocking the face, product branding, collage borders, duplicate limbs, neon cyberpunk clutter, malformed hands.初期生成で傘が二つ現れた場合は、そのまま使わず一つに修正します。マルチリファレンスは入力の誤りも拡大するため、見栄え以上に矛盾がないことが重要です。
画像3:空間参照から人物を完全に外す
場所の参照に別の人物がいると、Seedanceが第二の登場人物として解釈したり、その人物の位置までコピーしたりすることがあります。建築、奥行き、照明、天候、床の反射だけを示すクリーンプレートを用意します。

画像生成プロンプト
Use case: stylized-concept
Asset type: landscape reference image for a Seedance 2.0 multi-reference fashion-film workflow; environment and lighting anchor
Primary request: Create an original cinematic environment plate for a stylish fashion film, designed to define architecture, depth, atmosphere, reflections, and lighting without introducing a character.
Scene/backdrop: a vast rain-soaked brutalist underground transit concourse after closing, long concrete ribs and repeating rectangular portals, glossy black floor with shallow puddles, a distant opening filled with mist, sparse linear light fixtures.
Subject: environment only; no people. A restrained sequence of crimson light panels runs along one wall while cool cyan ceiling light and silver rain reflections establish the color script. One subtle wind current pushes mist and loose droplets through the corridor.
Style/medium: photorealistic cinematic location photography, premium fashion-film production design, original architecture.
Composition/framing: wide 16:9 landscape, strong one-point perspective down the concourse, low eye-level camera, 24mm lens character, clean central walking path, foreground puddle reflections and deep layered background.
Lighting/mood: cool cyan overhead pools, narrow crimson side accents, wet specular reflections, moody but readable, elegant tension.
Color palette: graphite concrete, black wet floor, cool cyan, restrained deep crimson, silver highlights.
Materials/textures: rough poured concrete, brushed metal edges, wet stone, fine rain mist, realistic puddle ripples.
Constraints: environment only, no humans or silhouettes; no text, letters, numbers, signage, logos, trademarks, watermarks, UI, vehicles, or copyrighted architecture; keep the central path unobstructed.
Avoid: neon cyberpunk city, colorful shop signs, crowded station, fantasy ruins, sci-fi spacecraft, excessive fog hiding the architecture, impossible reflections, fisheye distortion.中央の動線は人物が歩けるよう空けています。色の規則は、シアンの天井光、右側の赤い光、濡れた黒い床の三要素に絞りました。
画像4:キーフレームを参照間の契約書にする
最後の参照は、人物と空間を実際の開始画面のように合成したキーフレームです。人物の大きさ、カメラの高さ、レンズ感、画面内の位置、色調を決めます。物語全体を一枚に詰め込まず、動き始める直前の安定した状態を示します。

画像生成プロンプト
Use case: compositing
Asset type: landscape reference image for a Seedance 2.0 multi-reference fashion-film workflow; hero composition, lens, color-grade, and first-frame anchor
Input images: Image 1 is the wardrobe-and-prop reference—preserve the adult woman's face, short silver bob, crimson high-collar technical coat, matte black bodysuit, black gloves, chrome boots, and smoke-gray umbrella. Image 2 is the environment reference—preserve its rain-soaked brutalist transit concourse, one-point depth, wet black floor, cyan overhead light, and restrained crimson wall accents.
Primary request: Place the same performer from Image 1 naturally inside the central path of Image 2 and create a polished opening keyframe for a stylish fashion film.
Scene/backdrop: the same vast wet brutalist concourse after closing, fine rain drifting from the left openings, puddles reflecting cyan and crimson light.
Subject: the same clearly adult performer strides toward camera with composed confidence, holding the closed smoke-gray umbrella downward in her right hand. Her crimson coat hem lifts slightly in the crosswind; her face, hair, outfit proportions, and chrome boots remain recognizable.
Style/medium: photorealistic cinematic fashion editorial, original production design, realistic physical interaction.
Composition/framing: wide 16:9 landscape, low-angle medium-wide full-body shot, performer placed slightly left of center on the vanishing line, foreground puddle reflection, 28mm lens character, subtle natural motion in coat only, crisp face.
Lighting/mood: cool cyan overhead pools, narrow crimson edge light, silver rain highlights, elegant suspense and controlled energy.
Color palette: graphite, deep crimson, cool cyan, black, mirror silver.
Materials/textures: wet concrete, realistic puddles, matte coat fabric, chrome footwear, translucent smoke-gray umbrella.
Constraints: combine references without redesigning them; exactly one adult person and one closed umbrella; preserve character identity, wardrobe, prop proportions, architecture, and color script; physically plausible stance and reflections; no text, letters, numbers, signs, logos, trademarks, watermarks, UI, weapons, or copyrighted characters.
Avoid: face drift, different outfit, extra people, open umbrella, duplicate props, superhero pose, neon city clutter, extreme motion blur, floating feet, warped architecture, malformed hands.アップロード順だけでなく、プロンプトに役割を書く
DreaminaのMultiframesモードでは、アップロードした素材を@AssetNameで参照できます。実際の画面に表示される名前を使ってください。「画像1を参照」だけで終わらせず、引き継ぐ属性と引き継がない属性を明記します。
| 入力 | プロンプト内の役割 | 優先度 |
|---|---|---|
| @Image 1 | 人物、体格、髪、衣装全体のみ使用 | 最高 |
| @Image 2 | 傘、生地、手袋、金属素材のみ使用 | 高 |
| @Image 3 | 空間、照明、雨、反射のみ使用 | 高 |
| @Image 4 | 開始構図、レンズ、色調のみ使用 | 中 |
番号は実際のアップロード順に合わせます。動画の長さは文章内で指定するより生成画面の設定で選び、プロンプトは動く主体、カメラ経路、環境の反応に集中させます。
保存条件の後に動きを書く
良い動画プロンプトは、雰囲気の単語より出来事の順序が明確です。今回の例は次の六層で構成します。
- 各参照の役割と混ぜてはいけない要素
- 人物の開始姿勢、視線、最初の行動
- 歩行からターン、傘を開くまでの動作連鎖
- ロートラッキング、横移動、部分オービットのカメラ経路
- コート、髪、雨、水たまり、光の物理反応
- 維持条件、禁止要素、控えめなサウンド方向
そのまま使えるSeedance 2.0動画プロンプト
Use @Image 1 only for the adult performer's identity, short silver bob, facial features, body proportions, crimson high-collar technical coat, matte black bodysuit, black gloves, and chrome boots. Use @Image 2 only for the smoke-gray umbrella, coat seams, glove texture, and reflective chrome material details. Use @Image 3 only for the brutalist transit concourse, one-point depth, wet floor, cyan overhead lighting, restrained crimson wall accents, rain, mist, and reflection behavior. Use @Image 4 as the opening composition, lens language, scale, and color-grade reference. Do not merge or swap the roles of these references.
The same adult performer walks toward camera along the central path with calm confidence, holding the closed umbrella downward in her right hand. Begin with a low close tracking shot of one chrome boot stepping into a shallow puddle; water splashes naturally and the crimson coat edge passes through frame. Rise smoothly into a side-tracking medium-wide full-body shot as she continues walking. Her gaze stays forward, shoulders relaxed, coat hem and silver bob reacting consistently to the crosswind while rain strikes the floor and umbrella surface.
She slows, turns her head toward camera, then plants one foot and makes a controlled pivot. During the pivot she opens the single umbrella behind her shoulder in one physically plausible motion. The camera performs a restrained partial orbit in the opposite direction, preserving her face and body proportions. Crimson light passes through the translucent canopy, cyan highlights slide across the chrome boots, and reflected light moves across the wet floor. End on a stable low three-quarter hero frame with her calm gaze sharp, the open umbrella forming a clean circle behind her, and the corridor receding into mist.
Camera language: low macro tracking to medium-wide lateral tracking to controlled partial orbit; smooth acceleration and deceleration; no random cuts, no handheld shake, no extreme zoom, no impossible camera path.
Performance and physics: natural walking cadence, clear weight transfer, realistic coat drag and recovery, believable umbrella opening, coherent rain splash and reflections, stable hands, face, outfit, and prop proportions.
Audio direction: restrained industrial ambience, rain on concrete and umbrella fabric, precise chrome heel impacts, a soft coat swish, and one deep tonal pulse at the pivot; no dialogue and no dominant music.
Keep exactly one adult performer and one umbrella. Preserve identity, wardrobe, prop, architecture, color script, and lighting continuity across every shot. No extra people, duplicate objects, text, logos, signage, face drift, wardrobe changes, warped limbs, floating feet, impossible reflections, or overexposed highlights.長いプロンプトですが、各段落が異なる失敗を防ぎます。人物アンカーは顔と衣装、素材参照は小物、空間参照は照明と反射、キーフレームは開始カメラを担当し、テキストが動き、順序、優先度でつなぎます。
よくある失敗と修正方法
| 失敗 | 主な原因 | 最初に変える項目 |
|---|---|---|
| カットごとに顔が変わる | 顔が小さい、複数の人物が競合 | 人物アンカーを一人にする |
| 衣装や小物が複製される | 参照内に同じ物が複数ある | 一つだけ残す |
| 背景が別の都市になる | スタイル語が空間指定より強い | 空間の役割文を前へ移動 |
| カメラが暴れる | パン、ズーム、オービットが競合 | 動きを2〜3個に限定 |
| 歩行からターンが不自然 | 動作間の重心移動がない | 減速、足を置く、回転を順に書く |
| 静止画のように見える | 外見だけで環境反応がない | 髪、布、雨、反射の動きを追加 |
一つの結果が弱くても、すべてを書き直さないでください。人物が崩れたら人物参照と保存文だけ、カメラが崩れたらカメラ段落だけを修正します。シードやバリエーション機能が使える場合は一要素ずつ比較します。
生成前チェックリスト
- すべての参照で人物、小物、個数が矛盾していないか
- 人物アンカーで顔、手、全身、衣装の輪郭が読めるか
- 空間参照に不要な人物や文字がないか
- 各@Image行に引き継ぐ属性と除外する属性があるか
- 表情、視線、姿勢、行動、重心移動を説明したか
- カメラ経路が時系列でつながり、過剰ではないか
- 服、髪、雨、霧、水、反射が行動に反応するか
- 顔の変化、重複小物、文字、ロゴを禁止したか
- 実在人物を使う場合、本人の同意と必要な権利があるか
ByteDanceの公式発表では、実在人物の肖像を参照する場合、本人確認または事前の法的許可が必要になる場合があると案内しています。人物、ブランド、音楽、場所の素材は、自分で制作したものか利用権を持つものだけを使用してください。
代表サムネイルの生成プロンプト
次のプロンプトは、4枚の参照で設計した世界を一目で示すため、ターンと傘を開く瞬間をクライマックスとして制作したものです。
Use case: stylized-concept
Asset type: wide blog thumbnail and final-output concept for a guide about stylish Seedance 2.0 multi-reference video creation
Input images: Use the cinematic keyframe as the exact character, wardrobe, umbrella, environment, architecture, material, lighting, and color-grade reference.
Primary request: Create a striking original fashion-film climax that shows what the multi-reference package can produce while remaining recognizably connected to the input.
Scene/backdrop: the same rain-soaked brutalist transit concourse with cyan overhead light, restrained crimson wall accents, wet black floor, mist, and deep one-point perspective.
Subject: the same clearly adult silver-bob performer in the same crimson asymmetrical technical coat, matte black bodysuit, black gloves, and chrome boots. She has just pivoted sharply toward camera and opens the single translucent smoke-gray umbrella diagonally behind her shoulder; the coat arcs naturally with the turn, rain droplets sweep around the umbrella edge, and her calm gaze remains crisp and recognizable.
Style/medium: photorealistic cinematic luxury fashion campaign, energetic but physically plausible, original visual identity.
Composition/framing: wide 16:9 landscape, low three-quarter camera in a controlled partial orbit, performer large and slightly left of center, open umbrella forming a strong graphic circle behind her, long reflective corridor visible to the right, thumbnail-readable silhouette, clean negative space.
Lighting/mood: cool cyan overhead highlights, crimson rim light through the translucent umbrella, silver rain sparkle, elegant momentum and high-end editorial confidence.
Color palette: graphite black, deep crimson, cool cyan, smoke gray, mirror silver.
Materials/textures: wet concrete and puddles, matte technical coat fabric, translucent umbrella canopy, polished chrome boots, realistic rain droplets.
Constraints: preserve the exact adult character identity, silver bob, facial features, wardrobe, environment, and color script from the input; exactly one person and one open umbrella; realistic fabric and umbrella physics; crisp face; no text, letters, numbers, title, logos, trademarks, watermarks, UI, weapons, or copyrighted characters.
Avoid: identity drift, different clothing, extra people, duplicate umbrella, superhero effects, neon cyberpunk clutter, extreme blur, distorted umbrella spokes, floating feet, malformed hands, overexposed rain.まとめ:マルチリファレンスは枚数ではなく責任分離
Seedance 2.0の価値は多くのファイルを入れられることだけではなく、人物、空間、カメラ、動き、音を別々の資料から参照できる点にあります。人物、小物、空間、キーフレームを四つの明確な契約にし、プロンプトでも役割を宣言すれば、ビジュアルアイデンティティを保ちながら動きとカメラを変更しやすくなります。
主な出典
- ByteDance Seed — Seedance 2.0 Official Launch
- ByteDance Seed — Seedance 2.0 Model Page
- Dreamina — How to Use Seedance 2.0
- Dreamina — Consistent Characters, Styles and Scenes
サービス画面、参照上限、解像度、透かし、料金、商用利用条件はアカウントや地域によって異なる場合があります。制作前に公式サービスの最新ポリシーをご確認ください。