Transform one source video
Use the clip's motion, framing, timing, or performance as the foundation, then describe a new style, subject, environment, or treatment.
This AI video to video generator brings compatible models into one editor. Start with a source video when you want to preserve broad movement, pacing, or composition while changing the style or scene. On supported models, add images, clips, or audio to guide identity, motion, camera direction, and sound.
Use the clip's motion, framing, timing, or performance as the foundation, then describe a new style, subject, environment, or treatment.
Use separate images, videos, and supported audio for identity, appearance, motion, camera direction, or sound when the selected model allows it.
Start with source footage for a direct restyle, or give each optional reference a specific role and connect those roles in your prompt.
Upload footage as the motion foundation, then change its environment, lighting, materials, or overall visual style. The result is newly generated rather than a frame-perfect filter.
Example direction: Preserve the forward camera movement, but transform the street into a cinematic night scene with new lighting and architecture.Use a clear portrait or character image to guide identity, clothing, and visual details while your prompt directs the action, setting, and camera.
Example direction: Keep the character recognizable while introducing natural movement and a new environment.Combine visual references with a short performance clip when movement, rhythm, or camera timing matters. Assign each source a clear role in the prompt.
Example direction: Use the video for motion and pacing, then use the image reference for appearance.Bring character, product, style, motion, and supported audio references into one workflow to create a scene that no single source contains on its own.
Example direction: Give every uploaded item a numbered role and describe the final scene explicitly.The exact controls depend on the selected model and its supported source media.
Use an existing clip to guide composition, movement, pacing and performance while changing its visual treatment.
Turn a static portrait or character design into natural, expressive motion.
Use a reference performance to guide movement, rhythm and timing.
Inspect identity, texture, color, and motion across the full result because generative transformations can vary from the source.
Place familiar characters or products into completely new environments.
Combine references with a prompt to direct action, camera and atmosphere.
Choose the workflow that matches the media you already have.
Animate one still image, with an optional end frame on compatible models. This is the clearest choice when a single image should anchor the generated clip.
Animate an image →Begin from a written prompt when you do not need an existing image, clip, character, product, or audio source to guide the result.
Create from text →Transform existing footage when motion, pacing, composition, or performance should guide the result. Add references when identity, style, motion, and sound need separate sources of guidance.
Create a video →Begin with footage when its broad motion, timing, composition, or performance should guide a newly generated result. Multi-reference inputs are an optional secondary capability on compatible models.
| Capability | What to expect | Basis |
|---|---|---|
| Primary input | Source footage plus a transformation prompt | Workflow definition |
| Transformation guidance | Broad motion, pacing, composition, or performance can guide the output; it is not a frame-perfect filter | Provider guidance |
| Optional references | Images, additional video, or audio only when the selected model exposes those inputs | Live editor controls |
| Limits and credit estimate | Use the upload validation, output choices, and estimate shown for the current model | Live editor controls |
Common questions about creating and transforming videos with reference media.
An AI video to video generator uses an existing video as guidance for a newly generated clip. It can preserve broad motion, pacing, composition, or performance while a prompt changes the style, subject, lighting, environment, or creative direction.
Yes. Upload the source clip, choose a compatible model, and describe what should change and what should remain recognizable. The output is generative, so motion and composition are guidance rather than a frame-perfect copy of the original.
Use sharp, well-lit images where the character, product or subject is easy to identify. Front-facing portraits and unobstructed product views usually provide the clearest identity signal. When using several images, give each one a distinct purpose instead of uploading conflicting appearances.
A reference video can guide body movement, performance, camera direction, shot timing, composition and pacing. It acts as motion direction rather than a frame-perfect copy, so describe which movement or camera behavior matters in your prompt.
Yes, when the selected model supports those inputs. Images can guide identity or appearance, video can guide motion and camera behavior, and audio can provide sound or performance context. Audio support does not guarantee an exact voice match. The editor updates available media types and limits when you switch models.
State the role of each reference, then describe the new scene, action and camera direction. Include the details that must remain consistent and keep the main action clear. For multi-reference models, use the reference tags shown in the editor so the model can connect each instruction to the correct asset.
Video to Video starts with footage, so motion, timing, camera behavior, or performance can guide the result. Image to Video starts from a still frame and creates motion that was not present in the source. Use Text to Video when you want to begin without source media.
Yes, when the selected model supports it. You can combine several images, videos, or audio inputs and give each reference a different role in the generated scene. The editor updates the available inputs and limits for the selected model.
Limits depend on the selected model, including supported file types, dimensions, file size, video duration and the number of references. The editor validates uploads against those requirements. The credit estimate updates from the chosen model, resolution, output duration and any billable reference media.
Select a secure payment provider to continue.
Create amazing AI-generated videos and images
Create a new account with Google or GitHub to claim them. Unused credits expire 30 days after signup.
By continuing, you agree to the terms and privacy policy.