Veo3 AI
制作テクニック

一貫性の高いAI動画アセットシステムの作り方

人物、衣装、表情、動作、ロケーション、小道具、色彩の再利用可能な制作資料を作ります。

角度が変わると顔が変わる、動作中に服が再設計される、小道具の大きさが変わる、切り返しで別の部屋になる。多くの場合、足りないのは画質語ではなく視覚的な証拠です。

プロンプトは作業を指示し、アセットは事実と境界を示します。先に証拠を作り、その後でショットを演出します。

アセット固定するもの最小成果物継承しないもの
人物ID顔、髪、肌、体格正面、横顔、衣装正面、全身背面動作、場所、カメラ
衣装形、素材、色、靴ルックごとの全身資料表情と物語
表情感情と強度同じレンズと光の表情グリッド身体動作
動作経路、重心、接触、順序連続したキーポーズ人物IDと場所の様式
場所空間、光、素材、家具複数の無人アングルと平面図無関係な人物
小道具形、尺度、構造、素材前・横・後・詳細手と筋書き
色板色の役割と比率役割、HEX、用途物体の形

1. 最終プロンプトより先にアセットを作る理由

1枚の画像が証明できるのは、1つの角度、1つの光、1つの瞬間だけです。アセット一式は、カメラや衣装、感情、動作が変わっても維持すべき事実を示します。

各参照には制作上の役割を1つだけ与えます。

人物ID:正面、横顔、衣装正面、完全な背面
人物ID:正面、横顔、衣装正面、完全な背面
場所:連続するアングルと俯瞰平面図
場所:連続するアングルと俯瞰平面図
小道具:安定した形状と構造
小道具:安定した形状と構造
色板:役割、HEX、使用範囲
色板:役割、HEX、使用範囲

2. 人物アセット:IDを検証可能にする

まず清潔なマスターポートレートを作り、設定表へ広げます。横顔、顎、耳、後頭部、縫い目、靴、装飾、体格まで確認します。

顔、雰囲気、髪、初期衣装を固定するマスター画像
顔、雰囲気、髪、初期衣装を固定するマスター画像
正面、横顔、衣装正面、背面をまとめた人物設定表
正面、横顔、衣装正面、背面をまとめた人物設定表

テンプレートA — 顔写真と3つの全身ビュー

Create a photorealistic four-view character reference sheet from Image 1. Use a clean white studio background. On the left, show one highly detailed, perfectly front-facing live-action head-and-shoulders portrait. Preserve the exact facial structure, expression, hairstyle, hair accessories, pores, fine lines, eyelashes, flyaway hairs, and natural skin texture. No beauty retouching and no over-sharpening. On the right, show the exact same character in a standard A-pose as three aligned full-body views: front, side, and back. Preserve the same body proportions, complete wardrobe design, fabric layers, shoes, accessories, and realistic material behavior in every view. No scene, no action, no extra people, no logo, no watermark.

テンプレートB — 2つの顔アップと拡大衣装ビュー

Create a professional photorealistic character reference board from Image 1 on a neutral grey studio background. The layout must have two sections. The upper third contains two large face close-ups: a perfectly front-facing portrait on the left and a clean side-profile portrait on the right. Preserve the exact facial identity, bone structure, hair, makeup, age, skin texture, and accessories. The lower two-thirds contains two wardrobe views. On the left, show a magnified front-facing A-pose wardrobe view cropped from just below the neck to the shoes so the garment construction fills the frame. On the right, show a complete rear full-body view. Keep lighting, lens, body proportions, clothing, and materials identical across all panels. This is a production reference sheet, not a poster or fashion editorial. No action, no complex background, no logo, no watermark.

全身の一貫性には余白のある四面図、顔と衣装の細部には高密度テンプレートを使います。

3. 衣装アセット:ルックライブラリを作る

1枚の画像に人物IDと全衣装を同時に担当させません。中立な全身画像を作り、人物、ポーズ、カメラ、光を変えずに1ルックずつ移します。

衣装転送前の全身ベース
衣装転送前の全身ベース
人物参照と衣装参照を使った衣装転送
人物参照と衣装参照を使った衣装転送
Create a full-body studio image of the person in Image 1. Preserve the exact face, hairstyle, glasses, age, skin tone, body proportions, and wardrobe. Use a neutral grey background, standard relaxed A-pose, even studio lighting, and realistic anatomy. Keep the full body and shoes visible. No logo, watermark, or extra text.
Use Image 1 for the character's identity and body proportions. Use Image 2 only for the complete wardrobe, shoes, and accessories. Replace the outfit in Image 1 with the full outfit from Image 2 while preserving the same face, hairstyle, glasses, pose, body shape, lighting, camera, and neutral studio background. Do not change the person. No logo, watermark, or extra text.

承認した各衣装の正面・側面・背面を作り、LOOK_DAILY、LOOK_SPORT、LOOK_FORMAL と命名します。

衣装ライブラリ:DAILY、SPORT、FORMAL
衣装ライブラリ:DAILY、SPORT、FORMAL

4. 表情と動作のアセット

表情資料は感情による別人化を防ぎ、動作資料は開始から終了までの経路を維持します。

表情プロンプト

同一人物による16種類の抑制された表情
同一人物による16種類の抑制された表情
Using the exact same person from Image 1, create a photorealistic facial-expression asset sheet on a neutral grey studio background. Use a clean 4×4 grid with sixteen consistent head-and-shoulder portraits. Preserve the same facial structure, hairstyle, glasses, makeup, age, skin texture, lens, angle, and lighting in every cell. Show restrained, realistic variations: neutral, focused, soft smile, curious, surprised, tense, tired, calm, frown, angry, relieved, sad, thinking, happy, shy, and unbothered. Avoid cartoon acting, face drift, beauty retouching, and plastic skin. English labels only. No logo or watermark.

動作ビート用プロンプト

連続する16の動作ビート
連続する16の動作ビート
An unarmed female agent is ambushed by a larger attacker in a rain-soaked underground parking garage. She survives by evading, breaking a grip, using a restrained knee strike, sweeping, rolling, using a parked car as cover, disarming the attacker, and exiting. Convert this into a coherent 16-panel cinematic action-beat library. Each panel must show one readable action instant with a clear number, short English beat name, body direction, balance, contact point, and screen direction. Keep both character identities, wardrobe, garage geography, wet lighting, and camera language consistent. Modern grounded action-film tone, dark, cool, tense, realistic, non-graphic, no injury detail, no logo, no watermark.

物語に必要な抑制された表情を選びます。動作の各コマには、1つの瞬間、接触点、重心、画面方向だけを示します。

5. 場所アセット:シーンバイブルを作る

美しい1枚だけでは切り返しを証明できません。入口、逆方向、作業域、素材、常設小物、光、平面関係を記録します。

先に無人の空間を生成し、人物と物語動作は最終ショットに入れます。

複数の無人アングルと空間図によるシーンバイブル
複数の無人アングルと空間図によるシーンバイブル
Create a complete production scene bible for an empty modern community table-tennis hall used in a gentle slice-of-life comedy. The location must feel ordinary, authentic, slightly awkward, and lived-in rather than professional or spectacular. Use a clean 3×3 grid. Show: entrance wide shot; reverse wide shot; left-side view; right-side view; player eye-level view; low view along the net; bench and water station; close-up of table, paddle, balls and worn floor markings; and a readable top-down floor plan. Preserve the exact room geometry, table positions, windows, doors, lockers, benches, water dispenser, dark-green lower walls, warm-white upper walls, fluorescent lighting and scattered balls across every panel. No people, logos, subtitles, decorative poster design, or invented rooms. Use small English view labels only.

6. 小道具:カテゴリではなく形状を固定する

「黒いスポーツバッグ」や「赤いコーン」はカテゴリにすぎません。比率、縫い目、ファスナー、金具、素材、機能部分を資料で証明します。

前・横・後・機能詳細を示す小道具資料
前・横・後・機能詳細を示す小道具資料
Create a physical prop turnaround sheet on a white or light-grey background. Show the selected prop in front, side, and rear views, plus one close-up of its most important functional detail. Preserve the exact shape, color, material, scale, seams, hardware, and structural features across every view. Use clean English labels only. No hands, people, scene lighting, brand logo, or watermark.

7. 色板は色だけでなく役割を保存する

各色に役割と許容比率を設定します。環境の主色、衣装や素材の補助色、小面積のアクセントを区別します。

役割、HEX、用途を持つ制作色板
役割、HEX、用途を持つ制作色板
英語ラベルで素材の役割を指定したVeo3 AI画面
英語ラベルで素材の役割を指定したVeo3 AI画面
You are a film colorist, visual art director, and AI-video palette designer. Analyze all uploaded frame references as one visual system rather than as separate images. Extract exactly seven shared production colors: BASE, SUPPORT, SHADOW, HIGHLIGHT, SKIN, REFLECTION, and ACCENT. For each color, provide one #RRGGBB HEX value and one concise use case. Then create one clean professional palette asset image containing only the seven swatches, English role label, HEX value, and use case. Explain which colors may cover large areas, which must remain small accents, and which hues should be avoided. No people, scene illustration, long theory, logo, watermark, subtitles, neon saturation, or random colors.

タイムラインの前に各入力の担当を明記し、人物、動作、空間、光、色をモデルに推測させません。

8. アセット一式を動画モデルへ渡す

物語より先に割り当てを書きます。誰の素材か、何を参照するか、何を継承しないかを明確にします。

入力対象唯一の役割除外
画像1CHARACTER_A顔、髪、肌、体格スタジオ背景を無視
画像2LOOK_SPORT服、靴、装飾顔を再定義しない
画像3EXPRESSION_A感情と強度身体動作を変えない
画像4ACTION_A順序と接触人物IDと場所をコピーしない
画像5LOCATION_A空間、素材、家具、光偶然の物を無視
画像6PROP_A形、尺度、構造背景を無視
画像7PALETTE_A色の役割と比率肌を着色しない
REFERENCE ASSIGNMENTS
Image 1 defines CHARACTER_A's identity only: face, hairstyle, glasses, skin tone, and body proportions.
Image 2 defines LOOK_SPORT only: sports top, skirt, shoes, and accessories. Do not redefine the face.
Image 3 defines EXPRESSION_A only: focused but slightly embarrassed. Keep the performance restrained.
Image 4 defines ACTION_A only: the sequence, balance, contact points, and screen direction.
Image 5 defines LOCATION_A only: room geometry, tables, windows, doors, benches, and light direction.
Image 6 defines PROP_A only: shape, scale, material, seams, and hardware.
Image 7 defines PALETTE_A only: color roles and usage ratios.

Do not inherit unrelated backgrounds, people, text, logos, poses, or composition from any reference.

その後に時間軸、カメラ、音、連続性、除外条件を追加します。同じ素材の名前を途中で変えません。

生成前チェック

  • 常設人物ごとに1つのIDマスター
  • 顔を変えずに衣装変更できる
  • 表情強度に視覚的な上限がある
  • 動作に経路、重心、接触、終了状態がある
  • 場所に切り返しと平面図がある
  • 重要な小道具の形状が固定される
  • 各素材は1つの名前、対象、役割だけを持つ
  • 最終プロンプトに時間軸、連続性、音、除外がある

プロンプトは何をするかを伝え、アセットシステムはプロジェクトが何であるかを伝えます。事実が安定すれば、複雑なショットも演出と修正が容易になります。

On this page