Put the person from a still into a driving video. No pose estimation, no segmentation, no masks — the only preprocessing is repainting the clip's first frame. 4 sampling steps. Takes 4–10 s of driving video; anything longer is cut to 10.1 s.
No sign-in. The first-frame repaint is paid by this Space — 5 a day per visitor. The render runs on your own ZeroGPU quota and asks for more than a free day holds, so in practice PRO and up.
examples — click one for its driving clip, its painted first frame and the render
driving video — motion, framing and background are kept
…or frame 0, already painted (skips gpt-image-2)
step 2: render
What it produces
Every example above, rendered at 4 steps with seed 42. Each is a single paint and a single render; only the last three swap in a person.