One character, one action
A single performer turning or stepping gives you a clear movement to compare. Crowds and overlapping actions make it harder to tell whether the generated performance follows the intended reference.
Start a shot with a performance you can see. Video to Video AI uses a motion-reference video and a separate character image to generate a new character performance. Test one action, compare the result with the reference, and decide how it fits your next scene.
Build a character shotA gesture can establish a scene before you add a complicated setting. Use a greeting, a turn toward the camera or a few measured steps as the movement reference. Your character image supplies a visual starting point; the generated shot can reinterpret its face, outfit and surroundings.
Write the shot as a single moment rather than a sequence of events. Keep the camera direction compatible with the reference, and inspect the beginning, middle and end for changes in appearance. This workflow guides movement; it does not edit an existing film frame by frame.
The creative library below draws from this site’s published images and videos. Use it for composition and mood ideas; cards may represent other workflows. The reference-and-result comparisons in the generator are the motion-control examples.
Loading creative inspiration…
Choose a visible performance from the motion references, or sign in and upload an MP4 or MOV lasting 3–30 seconds and no larger than 32 MB. A continuous view of one performer is a practical starting point for a character shot.
Add a character image whose framing suits that performance. Use a full-body image for a walking shot and a closer image for a small gesture. Select an enabled Motion Control model and its available output quality.
Describe the setting and shot treatment, then review Required Credits for the current model, quality and billable duration. Generate the clip and inspect the completed performance in Your results before developing a longer scene.
For a character introduction, choose a motion with a clear start and a quiet finish. Look at whether the pose communicates the intended mood without needing dialogue. A restrained reference often makes the character easier to assess.
Treat separate generations as individual shots. Reusing an image is useful for testing a direction, but it does not guarantee continuity across clips. Review clothing, proportions and facial features before assembling a sequence elsewhere.
A single performer turning or stepping gives you a clear movement to compare. Crowds and overlapping actions make it harder to tell whether the generated performance follows the intended reference.
Keep the subject distinguishable from the background. Describe lighting and atmosphere without asking new objects to cross the character or obscure the main gesture.
Choose framing that shows the important body movement. Rapid camera instructions can compete with a steady reference, so begin with a simple shot and assess the result before adding complexity.
Watch the result without pausing and check whether its main action is understandable. Then compare its pace with the source performance using the paired videos.
Pause during turns and hand movements to look for shifting facial features, clothing or limbs. An attractive first frame does not establish consistency throughout the shot.
Keep the motion fixed while changing the image, or keep the image fixed while simplifying the reference. Comparing one change at a time helps you understand which input needs work.
This page needs a motion-reference video and a character image. Text adds scene direction, while the video guides the performance. Use a text-to-video workflow when you want to begin from a written scene alone.
You can use a result to explore a character action, framing or mood. It is a generated concept clip, so assess its details before using it to plan or communicate a larger sequence.
That is not guaranteed. Compare the face, costume and proportions in each completed clip, even when you reuse the same character image.
Treat the reference as movement guidance. This workflow does not promise copied music, spoken dialogue or lip synchronization. Check the finished clip and handle soundtrack work separately when needed.
Billable duration rounds up to whole seconds. The current model and output quality determine the quote shown before generation; review it again whenever you change an input or setting.
The reference already supplies the performance. Use the prompt to describe the scene and visual treatment, and avoid adding actions that conflict with the movement you selected.