Animate a photo or illustration
A still image supplies the subject and composition. Add an action that suits it: fabric moving in a breeze, a character turning toward the viewer, or a slow camera move around a product.
This image will be the starting frame of your video
0 / 5000
Generates video with AI audio (audio may be disabled for sensitive content)
Turn a photo into a video or generate a scene from text with Pixlr AI. Use image to video for product clips, animated artwork, and moving portraits, or start with a written scene when you do not have a reference.
Start with a visual you want to animate, or let a text prompt establish the look and action together.
A still image supplies the subject and composition. Add an action that suits it: fabric moving in a breeze, a character turning toward the viewer, or a slow camera move around a product.
Describe the location, subject, and event for text-to-video generation. This works well for an imagined establishing shot or supporting scene that does not need to match an existing picture.
Use the prompt to explain what happens during the clip, not just what the scene looks like.
Specify one main event, such as a hand lifting a cup or a character opening a book. For an uploaded image, spend more of the instruction on movement than on details already visible.
A fixed camera keeps attention on the subject’s action. A slow push-in emphasizes a detail; a sideways move reveals more of the scene. Describe the purpose of the move instead of combining conflicting directions.
Plan a brief clip around a gesture, reveal, or transition that can finish in the chosen time. If the idea involves several locations or separate events, generate individual shots for a longer edit.
Pixlr AI includes Wan 3.0, Seedance 2.5, and Kling 3.0. Select a model by the inputs, sound options, and output settings your scene needs.
Use an opening image to establish the shot. Where an end-frame input is available, add a destination image for the movement to lead toward. Keep both images compatible with the transition you want.
Supported reference modes accept images, video, or audio to guide a scene. Explain what each input contributes, such as a subject’s appearance, a movement, or a sound, rather than treating every reference as a fixed frame.
Choose a vertical, square, or landscape format for your intended placement where available. Set the clip length and resolution from the selected model’s controls before generating.
Use supported sound generation for audio that belongs to the action. For narration you want to revise independently, create a separate speech track and combine it with the clip in your editor.
Choose a shot that communicates one idea clearly, then review the full video before downloading it.
Use a restrained reveal or camera move to draw attention to a product. Watch its shape and markings throughout the clip, especially when the angle changes.
Give an illustration a small action or environmental movement that suits its style. Keep the focal character readable so motion adds interest without obscuring the design.
Create shots that illustrate a specific line or point in a script. Leave enough time to place the clip under the narration and trim the opening or ending in your editor.
Separate visual decisions from motion decisions when the subject or composition needs to be specific.
Use the image generator to develop the framing, lighting, and subject first. Choose a picture with enough space for the intended movement, then download it and upload it to the video workspace.
Add a motion instruction to that selected image. If the result changes an important face, object, or label, compare another take or simplify the action before building it into a finished sequence.
Photo animation, sound, references, and building a sequence from generated shots.
Choose image-to-video mode, upload a photo, and describe an action or camera movement. Select the available duration and output settings, then generate. Watch the whole clip before downloading so you can check how the subject changes over time.
Describe the movement you want to add: who or what moves, how the camera behaves, and how the moment ends. With a strong reference image, a focused motion instruction is usually more useful than repeating every visible color and object.
Yes. Text-to-video mode creates a scene from your description without requiring a starting photo. Include the subject, setting, action, and visual treatment. Use an image reference instead when matching a particular appearance is the priority.
Yes, with models and modes that support audio generation. Other supported workflows accept audio references. Choose the relevant controls for your scene; use a separate text-to-speech track if you need narration that can be revised independently.
Use an end frame when the clip needs to arrive at a specific final composition. A general reference guides an element of the scene without necessarily placing it at the end. These inputs serve different purposes and are available according to the selected model and mode.
Break the idea into individual shots and generate each one. Download the selected clips, arrange them in a video editor, and add narration, captions, and transitions there. Reuse suitable subject references to keep the sequence visually coherent.
Yes. Try free video generation from an image or a text prompt. A short scene with one clear action is a useful first test before developing a larger project.
Upload a picture or describe a scene, then direct its movement with Pixlr AI.
Start Free