Black-and-White Depth Video Driven · A Dress Showcase Workflow With Your Own Model

WanWorkflowFashionImage to Video7s
First convert a reference dance video into a black-and-white depth video, then generate your own model character card, and finally use a single-line prompt to have the model learn only the movement — not the original footage's person or scene — for a dress showcase against a pure white background. The original post is a three-step workflow, but the actual prompt fed to the model is just one sentence.
PROMPT · Video Prompt
Create with AI
Generate a women's fashion showcase video, using the person in reference image one, set against a pure white background; the video should only reference the person's dance movements, not the black-and-white footage, the person's identity, or the original scene.
✍️ Editor’s Notes

The prompt itself is thin — just one sentence — but the real cleverness is in splitting the workflow: pose/motion data and final-frame identity are fed to the model in two separate steps. The depth video first extracts a pure motion skeleton, stripping out the original footage's person and scene, so the third-step prompt only has to do one narrowing pass of 'positive inclusion plus negative exclusion.' Packing a whitelist and a blacklist into a single sentence — 'only reference A, not B or C' — is the cheapest way to keep an image-to-video model from drifting, and it works better than piling on shot details, because the motion is already locked in by the depth video; the prompt's only job is to name what shouldn't be learned.

Want tighter control? Editor-expanded reference version (not the original prompt):
Break it into four steps: 1. Framing: the person stands centered or turns slowly to show the full dress; keep the camera fixed or push in very slowly — don't follow the original video's camera angles. 2. Lighting: pure white background with soft, even studio lighting; fill light from all sides to remove shadows and bring out the fabric's texture and silhouette. 3. Motion pacing: match the timing and turning points of the reference dance, but the range of motion can be slightly reduced so the skirt's sway stays clear and doesn't blur. 4. Finish: hold the final beat on a pose that clearly shows the full-body silhouette, as the showcase's closing frame.
Create with AI
Workflow Prompt
PROMPT · Full Prompt
Create with AI
1. Find a reference video and generate a black-and-white depth video from it — this step is very fast, you can have codex generate it for you, or have it write you a python script to do it, either works. 2. Generate your own model — ideally with a character card, though a single image works too. 3. Prompt to generate the video: Generate a women's fashion showcase video, using the person in reference image one, set against a pure white background; the video should only reference the person's dance movements, not the black-and-white footage, the person's identity, or the original scene.
✍️ Editor’s Notes

This block is the three-step workflow itself, and the real division of labor happens in the first two steps: step one converts a reference video into a black-and-white depth map, keeping only the motion skeleton while depth information erases the original person and scene entirely — this step can even be handed off to codex to script and automate, no manual work needed. Step two separately generates your own model (a character card is preferred, though a single image works), completely decoupling 'who performs' from 'how they move.' Step three is the only part that's actually fed to the video model as a prompt, and it's just one sentence, because the first two steps have already locked down everything that needs locking. This layered approach — do the data prep inside the workflow, let the prompt only handle the finishing touch — is more stable than trying to cram every constraint into a single sentence.

Create with AI
Related Works
׋›
↑