Put the person from a still into a driving video. No pose estimation, no segmentation, no masks — the only preprocessing is repainting the clip's first frame. 4 sampling steps. Takes 4–10 s of driving video; anything longer is cut to 10.1 s.
Model card and weights · viggle.ai
Sign in to render — 5 a day per account. Rendering repaints your first frame with an image model this Space pays for.
Every example above, rendered at 4 steps with seed 42. Each is a single paint and a single render; only the last three swap in a person.