物体替换
16:9 10s Conversational edit
物体替换
AI 视频换物体、换背景提示词:对话式编辑三步走
用 Gemini Omni 的招牌功能对话式编辑,一轮一轮地换物体、换背景——人物和光线在多轮修改之间保持不变。
conversational-editing omni-flagship swap background iterative
提示词
Turn 1 — Generate base scene: A young woman in a white linen dress stands in a sunlit Victorian parlor, holding a porcelain teacup. Camera at chest height, soft window light from the left, locked-off shot, 10 seconds. Turn 2 — Swap one object: Change the porcelain teacup to a small bouquet of wildflowers. Keep her posture, the parlor, and the lighting exactly the same. Turn 3 — Replace the background: Now replace the parlor background with a misty forest clearing at dawn. Keep her, her dress, and the bouquet identical. Match the original key light direction (upper-left).
为什么这样写
对话式编辑(conversational editing)是 Gemini Omni 区别于 Sora / Veo 的招牌功能。Google 的 DeepMind 提示词指南专门演示过模型一轮一轮地替换物体(butterfly → bee → fireflies、astronaut → sea anemone),每一轮都在上一轮结果的基础上改。
这个三轮结构改编自 Atlas Cloud 的上手评测,他们验证的正是同一套流程:先生成基础画面 → 换一个变量 → 再换另一个。
来源等级:🟢 官方演示 + 媒体实测(高置信度)
万能句式:“Keep [X] exactly the same”
Seaart 的提示词合集和 Google 自己的文档都在强调:修改要求写得含糊,模型就会越改越多、画面走样。换一样东西时,明确列出你想保留的东西。
“Replace the food on the plates with creamy pumpkin soup. Keep the two people talking, their movements, facial expressions, ocean background, and all lighting exactly the same.”
(把盘子里的食物换成奶油南瓜汤。正在交谈的两个人、他们的动作、表情、海景背景和所有灯光都保持完全不变。)
—— Seaart 提示词合集
怎么改
- 单变量原则:每轮只改一样东西。一条提示词里同时改好几处,Omni 会乱。
- 锁定清单:明确写出要保留的脸 / 服装 / 姿势 / 灯光 / 背景
- 一环扣一环:每一轮针对的都是上一轮的结果(而不是最初那一版)
- 重来的时机:如果改了 4~5 轮后走样越来越明显,就从一个新的基础画面重新生成
常见翻车点
据 TechCrunch 的发布报道和 Atlas Cloud 的测试:
- 不写“Keep X identical”的话,Omni 换背景时可能顺手把人物的衣服或脸也改了
- 一轮里换好几个物体,走样的方向完全没法预料
- 超过 4 轮后,面部细节和手部动作都会变差(Atlas Cloud 给角色一致性打了 3/5)
备注
- 多轮编辑需要订阅 Gemini Omni Flash(在 Gemini App 或 Google Flow 里使用)
- 所有成片都带 Google SynthID 隐形水印——设计如此,去不掉
- 单条片段最长 10 秒(Flash 的硬上限,TechCrunch 和 CineD 都确认过)
来源
相关