Multi-Reference Womenswear Fashion Display with Depth Motion Control
This prompt generates a womenswear showcase video by combining three references: an identity/outfit image, a background scene image, and a black-and-white depth video for motion. The character from image one performs the dance/pose motion from the depth video while naturally integrated into the scene from image two, with strict identity, clothing, and background consistency.
Prompt (original, unmodified)
生成一段女装展示视频。 人物身份与服装严格参考图片一,保持人物脸型、五官、发型、身材比例和服装款式一致。 场景与环境严格参考图片二,将图片二作为背景参考,尽量保持其空间结构、布景、色调、光线方向和整体氛围一致,使人物自然融入该场景。 动作参考提供的黑白深度视频,仅提取其中人物的身体姿态、舞蹈动作、动作节奏、肢体轨迹和镜头中的运动关系。 黑白深度视频只作为动作控制参考,不要参考其中的黑白视觉效果、人物身份、人物外貌、服装、背景、原始场景、材质或色彩。 最终视频应表现为:图片一中的人物,在图片二的真实场景中,按照深度视频中的动作自然进行女装展示。保持人物身份、服装 and 背景稳定,动作流畅自然,避免脸部漂移、服装变化、背景变化、肢体畸变和人物闪烁。
Reproduced verbatim from the original post. Never edited or translated. Everything outside this block is our own commentary.