Seedance 2.0 完全プロンプトガイド
Seedance 2.0 のテキスト動画生成、画像動画生成、参照素材からの動画生成に効果的なプロンプトの書き方を解説します。
Seedance 2.0 は、クリエイティブな動画生成に対応するマルチモーダルモデルです。テキスト、画像、動画、音声の参照素材から、意図のある動き、カメラ表現、映像と音の雰囲気を備えた動画を生成できます。このガイドでは、適切な生成モードの選び方と、明確で実行可能なプロンプトを通じて制作意図を伝える方法を解説します。
1. Seedance 2.0 でできること
1.1 テキストから動画を生成
テキスト動画生成では、テキストだけをもとにシーンを構築します。被写体、動作、環境、スタイル、カメラ、音を記述すると、モデルがそのコンセプトを完全な動くシーンへと発展させます。
このモードは最も自由度が高く、コンセプトのプレビュー、広告アイデア、ムード映像、幻想的なシーン、まだビジュアル素材がない段階での素早い検討に適しています。視覚的な基準となる画像や動画がないため、プロンプトの具体性が被写体のデザイン、空間関係、動きに直接影響します。
1.2 画像から動画を生成
画像動画生成では、静止写真、イラスト、商品画像、コンセプトアートをアニメーション化します。入力画像によって被写体の外観、構図、カラーパレット、基本的な環境はすでに決まっています。そのため、プロンプトで画像を単に説明し直すのではなく、次に何が起きるか、シーンがどう動くかを伝えます。
人物の動き、微表情、髪や服の物理的な動き、風・霧・光などの環境効果、プッシュイン・パン・固定ショットなどのカメラ動作を制御できます。イラスト、水彩画、3D アートを動かす場合は、元の画材表現を保ち、別のビジュアルスタイルへ変化しないようモデルに指示することもできます。
1.3 参照素材から動画を生成
参照動画生成は、より強い制御と一貫性が必要な制作向けです。画像、動画、音声を参照素材として与え、それぞれに結果の特定要素を担当させられます。たとえば、人物の外観、商品のディテール、動き、カメラワーク、アートスタイル、リズム、感情の展開などです。
重要なのは、各素材の役割を明示することです。たとえば、@image1 で人物の外観を定義し、@video1 からは動きとカメラのリズムだけを取り入れ、@audio1 で感情の展開を動かせます。参照素材が増えるほど、外観・動き・スタイルの衝突を防ぐため、明確な優先順位が重要になります。
1.4 適切な生成モードの選び方
- アイデアと文章による説明だけがある場合: テキスト動画生成を選び、被写体、シーン、カメラ演出をゼロから構築します。
- 優れた最初のフレーム、ポートレート、商品画像、アート作品がすでにある場合: 画像動画生成を選び、既存の画像をどう動かすかにプロンプトを集中させます。
- 人物の外観、動き、カメラ表現、スタイル、音声のリズムを再利用したい場合: 参照動画生成を選び、すべての参照素材に具体的な目的を割り当てます。
2. 効果的なプロンプトの基礎
Seedance 2.0 のプロンプトを書く際は、次の原則を活用してください。
- 具体的かつ明確に書く: 被写体の外観、動き、環境、照明、音など、目や耳で確認できるディテールを記述します。「美しい」や「高級感のある」のような曖昧な言葉は、実際のシーンの見え方を説明する具体的な内容に置き換えます。
- 論理的な順序で整理する: 被写体、動作、環境、スタイル、カメラ、音の順に並べると安定します。
便利なプロンプト構成は次のとおりです。
| 要素 | 記述する内容 | 例 |
|---|---|---|
| 被写体 / 人物 | 主な被写体、外観、衣装、重要な小物 | 深緑色のコートを着て、使い込まれた木製のバイオリンを持つ若いバイオリニスト |
| 動作 / 動き | 被写体が何をするか。速度、方向、リズムも含む | 駅の中をゆっくり歩き、時計の下で立ち止まって演奏を始める |
| 環境 / 舞台 | 場所、時間、天候、空間の奥行き、背景の動き | 夜明けの誰もいない古い鉄道駅。線路の上を薄い霧が流れている |
| スタイル / ムード | 映像表現、色、照明、感情的なトーン | 映画的なリアリズム、静かで希望に満ちた雰囲気、冷たい青の影と柔らかな金色の朝日 |
| カメラ / 構図 | ショットサイズ、アングル、カメラ移動、フォーカス、フレーミング | ワイドショットから始め、浅い被写界深度のミディアムクローズアップへゆっくりドリーインする |
| 音声 / サウンド | 環境音、動作音、音楽、セリフ | 遠くの列車の環境音、柔らかな風、明瞭なバイオリン独奏、セリフなし |
1 つの完全なプロンプトにまとめると、次のようになります。
A young violinist in a dark green coat, carrying a weathered wooden violin, walks slowly through an empty old railway station at dawn, then stops and begins to play beneath the clock. Light fog drifts across the tracks. Cinematic realism, quiet and hopeful, with cool blue shadows and soft golden morning light. Begin with a wide shot, then make a slow dolly-in to a medium close-up with shallow depth of field. Distant train ambience, soft wind, clear solo violin, no dialogue.- 正確なキーワードを使う:
soft lighting、watercolor style、slow tracking shotのような表現で、質感、アートスタイル、カメラ表現を定義します。重複する言葉や矛盾する言葉を積み重ねないようにします。 - 重要な除外条件を書く: 「no cuts」「do not change the character's appearance」「do not make the image photorealistic」のような否定指示を、重要な制約に使用します。本当に必要な制約だけを残してください。
- 参照素材の役割を明確に割り当てる: 画像、動画、音声ファイルをアップロードしたら、外観、動き、カメラ動作、音楽のどれを制御するか指定します。
- テストして反復する: 最初の結果が完璧とは限りません。一度に変更する語句、ディテール、優先順位を少数に絞り、どの指示が出力を改善したのか確認できるようにします。
3. Seedance 2.0 のプロンプトフレームワーク
3.1 テキスト動画生成
テキスト動画生成は、Seedance 2.0 で最も直接的なワークフローです。視覚的な基準となる画像や動画がないため、プロンプトの構成と具体性が特に重要です。
プロンプト
A solitary meteorologist in a bright orange weather suit stands on a black volcanic ridge as a vast thunderstorm approaches across the ocean. She raises a handheld sensor into the wind; her coat and loose straps whip violently while sheets of rain sweep across the rocks. A distant lightning strike illuminates the cloud layers, and she turns toward the flash with a focused expression. Photorealistic cinematic drama, cold steel-blue palette, wet reflective textures, strong backlight through the rain. Start with a wide establishing shot, then track slowly around her to a low-angle medium shot as the lightning flashes. Deep wind, rolling thunder, rain striking fabric and stone, no dialogue.出力 · 動画
- 被写体から書き始める: 最も重要な人物、商品、物体をプロンプトの冒頭に置くと、モデルが視覚的な焦点を素早く確立できます。
- 動きを明確に記述する: 動きは動画の核です。
slowly panning left、rapidly spinning、gently swayingのような表現で、動きの種類、速度、方向を指定します。 - テンポを定義する: タイミングが重要な場合は、
a slow, meditative sequenceやa fast-paced, high-energy montageのようにシーケンスを表現します。 - 照明を定義する:
soft, diffused morning light、harsh neon backlighting、flickering candlelight casting warm shadowsのように、光源、方向、光の質を説明します。 - 雰囲気を加える:
tension-filled、whimsical、melancholic、euphoricなどの語句は、感情的なトーンを定めるのに役立ちます。
3.2 画像動画生成
静止写真、アート作品、コンセプト画像を与えると、プロンプトによって Seedance 2.0 に既存のフレームをどう動かすかを伝えられます。
テキスト動画生成とは異なり、シーンを文章で一から作り直す必要はありません。次に何が起きるか、画像の各レイヤーをどう動かすかに集中します。
レイヤーごとに動きを設計する
プロンプト
Animate with fine red dust sweeping across the cracked ground in the foreground. The astronaut's loose fabric straps flutter gently in a steady wind, and faint condensation gathers along the inside edge of the visor. In the background, red warning lights flicker on the abandoned outpost, and two reconnaissance drones circle lazily above the antenna towers. Apply a slow, contemplative camera push-in toward the astronaut. The mood is lonely and monumental, with a muted rust-and-cyan color grade and soft, diffused sunset lighting.入力 · @image1

出力 · 動画
効果的な画像動画プロンプトでは、シーンを独立して動かせる複数のモーションレイヤーとして捉えます。
初心者によくある失敗は、環境の変化を無視して被写体の動きだけを記述することです。前景、被写体、背景の動きを連携させると、通常はより自然で映画的な結果になります。
- 前景: 動く葉、揺らめくろうそくの光、水面の波紋など、カメラに最も近い要素。
- 中景: ゆっくり顔を向ける人物や、重心を移す馬など、主な被写体と中心的な動作。
- 背景: 流れる雲、遠くの旗、被写体のはるか後方を歩く人々など、奥行きを生み出す要素。
3 つのレイヤーを同じ強さで動かす必要はありません。被写体の動作を明確に保ち、前景と背景のより繊細な動きで支えます。
微細な動きで感情を伝える
人物や顔のクローズアップを含む画像では、表情や姿勢のわずかな変化によって、不安定な動きを加えることなく被写体に生命感を与えられます。
小さな視線移動、わずかに目を細める動き、見える呼吸、風で動く襟などは、大げさな演技より多くの感情を伝えることがあります。
プロンプト
Animate with a barely perceptible shift in the fisherman's gaze—his eyes slowly tracking something distant on the horizon. A faint squint tightens around his eyes. His jacket collar flutters softly. Waves reflect subtly in his eyes. The camera remains completely static, locked off. The atmosphere is deeply contemplative and nostalgic, desaturated with warm tones with soft coastal light.入力 · @image1

出力 · 動画
抽象的・芸術的なスタイルを保つ
画像動画生成では、写真だけでなく、絵画、イラスト、コンセプトアートも動かせます。
こうした参照素材では、元の画材が持つ筆致、輪郭、色、質感、素材感を保つよう明示してください。これにより、アニメーションが過度に写実的になったり、視覚的な一貫性を失ったりすることを防げます。
プロンプト
Animate this scene while fully preserving the dreamy watercolor storybook aesthetic—soft, luminous color washes, delicate painterly textures, glowing gold details, and diffused starlight throughout. The little boy gently plays the harp, his fingers softly plucking the strings as they shimmer and vibrate with golden light. White birds circle gracefully around the harp, while swallows glide and flutter across the starry sky in smooth, flowing paths. The animation should feel hand-crafted, poetic, and delicate, never sharp or digital. A gentle, whimsical atmosphere with celestial blue, soft white, and warm golden tones.入力 · @image1

出力 · 動画
画像動画生成のキーワード集
固定の公式ではありません。必要な動きと雰囲気に合わせて組み合わせてください。
| カテゴリー | 有用なキーワード |
|---|---|
| 動きの質 | gently、barely perceptible、slowly drifting、rhythmically swaying、subtly rippling |
| 雰囲気 | mist rolling in、particles of dust、heat haze、soft bokeh、volumetric light rays |
| 人物の生命感 | micro-expression shift、eyes slowly tracking、breath visible、hair softly lifted by wind |
| カメラ | locked off、slow push-in、subtle drift、gentle handheld sway、rack focus |
| スタイルの維持 | maintain painterly texture、preserve film grain、honor the original color palette |
3.3 参照動画生成
「維持」と「変換」の手法
参照動画プロンプトは、維持と変換という 2 つの明確な部分に分けられます。これにより、何を一貫させ、何を変えるべきかをモデル任せにせず、Seedance 2.0 に直接伝えられます。
プロンプト
(Preserve) Retain all original movement, choreography, timing, and body posture of the dancer exactly as they appear in the source video. Maintain the original camera angle and framing throughout.
(Transform) Re-stylize the entire visual environment as an ethereal, otherworldly forest glade. Replace the studio floor with a carpet of luminous, floating flower petals. Surround the dancer with slow-moving fireflies and drifting luminescent spores. The dancer's costume should transform into a flowing, translucent gown that trails light. Apply a dreamlike, fantasy aesthetic with soft teal and lavender tones, volumetric god rays filtering through ancient trees. Film grain texture, cinematic quality.入力 · @video1
出力 · 動画
スタイル変換:新しいビジュアル言語を定義する
元動画の内容を保ちながら全体のスタイルを変えたい場合、広いスタイル名だけを指定するのではなく、目的とする美学を具体的に説明します。パレット、照明、素材、画面の質感、時代を示すディテールを定義してください。
また、被写体の動作、演技のタイミング、セリフのタイミング、カメラ移動など、元動画のどの要素を維持するかも明記します。
プロンプト
Preserve all dialogue timing, gestures, and the camera position. Re-stylize the entire scene in the visual language of a Studio Ghibli animated feature—soft, hand-drawn cel animation aesthetic, warm and richly textured backgrounds, characters rendered with expressive Ghibli-style proportions. The café transforms into a charming, vintage European bakery with afternoon sunlight streaming through lace curtains. Palette is warm, creamy, and inviting. Gentle ambient sounds of clinking cups and soft piano music implied in the visual atmosphere.入力 · @video1
出力 · 動画
マルチモーダル融合
マルチモーダル生成では、画像、動画、音声の参照素材を 1 回のリクエストで組み合わせられます。一方で自由度が高まるほど複雑にもなります。各素材が異なるスタイル、リズム、カラーパレット、感情的なトーンを持ち込む可能性があるためです。目標は、統一された制作方針を確立し、参照素材同士が競合しないようにすることです。
制作上の優先順位を決める
入力素材を映画制作の各部門として考えてみましょう。ある参照素材がビジュアルアイデンティティを定義し、別の素材が動きとカメラ動作を振り付け、さらに別の素材が感情のテンポを制御します。それぞれの素材に、明確な役割を 1 つ割り当てます。
プロンプト
@image1 is the primary visual authority—the protagonist’s face, hairstyle, blue hair streak, cybernetic eye implant, black tactical clothing, illuminated cyan details, boots, and messenger bag must remain exactly consistent throughout the entire video. Do not redesign, replace, or simplify any part of her appearance.
@video1 serves exclusively as the movement, parkour choreography, body-mechanics, and camera reference. Apply its exact sprinting rhythm, barrier vault, wall run, landing impact, low slide, and final acceleration to the protagonist from @image1. Preserve the original sequence, timing, spatial direction, tracking shots, camera orbit, and low floor-level camera movement, but do not carry over the gray training outfit, stunt performer’s identity, warehouse, or any other visual element from @video1.
@audio1 sets the emotional rhythm and editing intensity of the entire sequence. During the restrained opening pulses, begin with controlled running and a smooth low-angle tracking shot. As the percussion builds, increase the protagonist’s speed, environmental motion, and camera energy. Synchronize the vault, wall push, slide, and strongest camera movements with the major rhythmic accents. At the musical drop, reveal the chase at full intensity and finish on the final impact.入力 · @image1

入力 · @video1
入力 · @audio1
出力 · 動画
融合手法:同じ重みの参照素材
2 つ以上の素材を本当に融合して新しいビジュアル世界を作る場合は、どれか 1 つを優先すると宣言するのではなく、それらをどう組み合わせるか説明します。
プロンプト
Fuse the visual identities of @image1 and @image2 equally into a single, cohesive world—a retro-futurist city that exists at the intersection of 1930s art deco grandeur and contemporary neon Tokyo nightlife. Neither should dominate; the architecture carries the geometric elegance of @image2 while glowing with the saturated neon palette and wet-reflective streets of @image1. Animate a slow, gliding aerial camera drift through this world, unhurried and contemplative. Let @audio1 dictate the pace entirely—every camera movement should feel as languid and swinging as the jazz rhythm. The atmosphere is nostalgic, mysterious, and quietly beautiful.入力 · @audio1
入力 · @image1

入力 · @image2

出力 · 動画
音声を主な駆動要素として使う
音楽とサウンドデザインは、動画の構成を最初から最後まで制御できます。音の変化に合わせてシーンがどう反応するか説明してください。静かな部分では動きを抑え、音楽が盛り上がるにつれて環境の動きを増やし、クライマックスで最も強い視覚的変化を起こします。
プロンプト
Let @audio1 be the architect of this entire video. Begin in near-silence: a static, locked-off shot of the lighthouse from @image1—still, barely animated, only the faintest movement of stormy clouds. As the orchestral score begins to swell, incrementally increase the intensity of the environment—waves grow larger, lightning begins to flash in the distance, the wind picks up, the lighthouse beam begins to rotate. By the time the score reaches its full crescendo, the scene should be a breathtaking storm in full fury—crashing waves, torrential rain, dramatic lightning strikes illuminating the cliff face, the lighthouse beam cutting through the chaos. The visuals and music must feel inseparable, as if one created the other. Cinematic, photorealistic, deeply dramatic.入力 · @audio1
入力 · @image1

出力 · 動画
4. まとめ
優れたプロンプトは、形容詞を積み重ねたものではありません。誰または何が被写体なのか、何が起きるのか、環境がどう反応するのか、カメラがシーンをどう捉えるのか、音が感情の展開をどう形作るのかという、制作上の役割分担を明確にしたものです。
テキスト動画生成はゼロから作り出します。画像動画生成は既存のフレームに生命を吹き込みます。参照動画生成は、素材ごとの明確な役割によって外観、動き、スタイル、リズムを固定します。
まず 1 つの明確なビジュアルアイデアから始め、動き、カメラ動作、雰囲気を少しずつ調整してください。一度にすべてを指定しようとするよりも、意図を明確にするほうが安定した結果につながります。