All tools

Clone yourself on video with AI

Your phone clip goes to three Wan 3.0 nodes that multiply you: five copies repeating your gesture a beat late, copies at every scale, a whole crowd of you.

Interactive demo

Real workflow, real results. Pan and zoom freely.

Use this workflow

The real workflow: the starting clip wired as a reference to three Wan 3.0 Video nodes, one cloning treatment per node. The three videos are the demo's own, the same face on every copy.

At a glance

What the workflow produces
3 videos of 15s in 720p, 9:16
Canvas structure
8 nodes wired by 6 connections
Models wired in
Wan 3.0
Cost of one full run
about 180 credits, roughly $2.16

Video cloning used to need a tripod, a fixed background, several takes and one mask per character. Here it needs a five second clip filmed by hand: you clap, you take a step, and the model fills the frame with you.

The three treatments in the demo do not clone the same way: five copies in an arc that pick your gesture up one after the other, like a wave; copies at every size sliding past one another; and a whole crowd, packed shoulder to shoulder to the back wall.

The hard part is not the number of copies, it is their FIDELITY: the moment one copy comes out with another face or another jacket, the effect collapses. So the three prompts repeat that every copy is the same person. Wan 3.0 takes your clip as a reference and keeps the face, the clothes and the gesture. The three videos cost one hundred and eighty credits.

How it works

  1. 1

    Film somewhere open

    A hall, a car park, an empty room, a plain background. The clones need space, and a busy set hides half of them behind the furniture.

  2. 2

    One short clean gesture

    A clap, a step sideways, a wave. A simple gesture repeats cleanly from one copy to the next, where a sentence or a choreography drifts out of sync.

  3. 3

    Keep the fidelity sentence

    Every prompt says that all the copies are the same person, same clothes included. That is what stops the model from filling the frame with strangers who vaguely look like you.

  4. 4

    Generate all three, compare

    "Generate all" outputs the three treatments. The delayed one tells an idea, the crowd makes the effect, the scale makes the picture: the choice depends on what you are telling.

Variations and use cases

The same canvas, tuned for different needs: each variation is two or three lines of prompt away.

The delayed wave

Five copies in an arc picking your gesture up one after the other, a beat late. It is the most readable treatment, and the one that works best on a simple gesture.

Copies at every scale

Huge versions of you in the foreground, tiny ones at the back, some at knee height, sliding past one another. A graphic render, close to a fashion edit.

The whole crowd

Dozens of you packed from the front of the frame to the wall, clapping at the same instant. It is the most spectacular, and the hardest to hold together on the likeness.

The argument with yourself

Two copies face to face answering each other, rendered from a clip of you talking. A short video format that travels well, and it needs a very still gesture at the start.

The one person team

A freelancer multiplied into a meeting, a shoot, a customer service desk. The picture says in three seconds what a paragraph explains badly on an about page.

The synchronised choreography

One dance step picked up by the whole crowd at the same instant. Keep the movement short: the longer it runs, the more the copies at the back end up going their own way.

Common mistakes

Filming in a furnished room

A sofa, a table and a shelf eat the space where the clones should be standing, and the model puts two of them instead of ten. A bare wall and clear floor beat a nice set.

Making too long a gesture

A five second choreography drifts out of sync between the copies and gives a crowd fidgeting for no reason. Two seconds of clear gesture, held, are worth more.

Forgetting to say it is the same person

Without that sentence, the model fills the frame with men and women who look like you from a distance, and the cloning effect gives way to an ordinary crowd.

Wearing very detailed clothes

A complicated pattern or a precise logo degrades on the distant copies. A plain outfit holds the repetition far better, and that is what the demo does.

Frequently asked questions

How many copies can I have?

The crowd in the demo counts several dozen. Beyond that the model loses the likeness on the copies at the back, which barely shows in vertical but shows on a large screen.

Do the copies move together?

That is something you ask for. The first treatment deliberately offsets them by a fraction of a second to build a wave, the third has them all clap at the same instant.

Can I clone two different people?

Film them both in the starting clip and ask for each group of copies to keep the right person. It is more fragile than with one, expect several attempts.

How much does one cloning cost?

One hundred and eighty credits for the three treatments, sixty per five second video in 720p with sound. Rerunning a single treatment costs only sixty credits.

Does the sound follow the copies?

The node generates the audio with the video, and the crowd claps in a single sound. For a precise render, replace the sound in the edit.

Can I string the three treatments together?

Yes, the Edit node assembles them in the browser with no extra credit. Fifteen seconds of cloning make a video that stands on its own.

How long should the starting clip be?

Between one and fifteen seconds. Five are enough, and a shorter clip leaves the model fewer chances to lose the likeness.

Does it work with an animal?

Yes, the mechanism does not assume a human. Film your dog in an empty corridor and ask for faithful copies: the likeness lock applies in exactly the same way.

Clone yourself on video with AI

Use this workflow

Free account, no card required. The workflow opens pre-filled in your canvas.