マルチモーダル参照シネマティック・VFX
SFミステリー予告編
2枚の参照画像で雰囲気と主人公を固定し、プロンプトでプッシュイン、タイトル、音響を制御する予告編です。
詳細を表示
完全なプロンプト(検証済み英語原文)
Realistic cinematic look, high-contrast lighting, and a tight pace. Use Figure 1 as the overall atmosphere and style reference, and Figure 2 as the protagonist reference. Shot 1 — Ultra-wide establishing shot. A huge circular cosmic gateway nearly fills the frame. The person is only a tiny figure seen from behind before the gateway, positioned toward the lower right. The ground is wet and reflective, and the center of the gateway is pitch black. The camera slowly pushes forward. A large title fades in from the edge of the darkness, blurred at first and then sharp: "THE STARS WERE LISTENING". Use an extremely condensed, heavy, all-caps typeface in dark red mixed with rust red, with subtle grain and misted edges. Audio: a deep low-frequency pulse, faint metallic vibrations in the distance, and a soft hit as the text becomes sharp. → Hard cut.
必要な入力
- 画像1:雰囲気とスタイルの基準。カットに引き継ぎたい環境、配色、カラーグレーディングを含めます。
- 画像2:主人公の参照。人物が画面内で小さくなっても顔とシルエットが判別できる大きさで撮影します。
- 実際に焼き込みたいタイトル文字。H3 は書いたとおりに文字を描画しようとするため、短くし綴りを正確にします。
機能する理由
- 参照素材ごとに役割が1つだけ与えられています(画像1が雰囲気、画像2が人物)。どちらがルックを決めるかをモデルが推測せずに済みます。
- カットは1回の連続プッシュインと1つの出来事(タイトルが鮮明になる)だけで構成されており、15秒が実際に支えられる情報量に収まっています。
- 音は「映画的な劇伴」といった曖昧な指定ではなく、低域のパルス、遠くの金属的な振動、文字が鮮明になる瞬間の打撃音という3層で書かれています。
置き換え可能な変数
- タイトル文字と書体処理
- 設定ショットのゲートやランドマーク
- タイトルの色
- 環境音のベッド
- 終わりのカットの扱い
制約
- 素材は配列上の位置で指定します(「Figure 1」「Figure 2」)。image_urls の並び順と一致させてください。@image1 という記法はこの API 契約には含まれません。
- 公開されているプロンプトは意図的に途中で止めた出発点です。ハードカットで終わるため、2つ目のカットを自分で続けられます。
設定
マルチモーダル参照 · 15s · 16:9 · 参照素材 2件










