OpenAI Sora 2
Use an image as the visual starting point and describe the motion you want to see. The live selector shows the output settings available for the chosen Sora 2 version instead of assuming every version has the same contract.
No videos generated yet. Upload an image to start!
No videos generated yet. Upload an image to start!
Every model on this page accepts an opening image, but its duration, resolution, aspect ratio and end-frame support can differ. Choose the version after deciding whether you need a simple animated photo or a transition between two defined frames.
Use an image as the visual starting point and describe the motion you want to see. The live selector shows the output settings available for the chosen Sora 2 version instead of assuming every version has the same contract.
Veo variants support frame-guided generation through the image workflow. Versions that allow a final frame expose it directly, so you can define an opening composition and, when supported, a destination composition.
MiniMax Hailuo H3 can generate from keyframes with its own resolution and duration choices. On this page the form stays in frame mode, so unrelated video or audio references cannot enter the request.
Seedance and Wan versions provide image-to-video options for short clips. Select the exact version first, then review the visible frame controls and cost before submitting the image and prompt.
Start with a usable frame, explain the movement rather than repeating everything visible, and review the model-specific output choices before generation.
Choose a clear photo, illustration, render or product image. The first frame anchors the subject and composition, so crop it intentionally and remove distracting details before upload. Supported models may also let you add a final frame.
Compare the model versions in the selector. Changing the model also changes the legal aspect ratios, resolutions and durations; an end-frame control appears only where the selected version supports that input.
Write one observable shot. State how the subject moves, what the camera does and which visual details should remain stable. Avoid asking for several unrelated scenes inside one short generation.
Choose from the settings shown for the current model, check the calculated credit amount and submit. The task appears in Recent Videos, where the completed clip can be previewed, downloaded or used for another attempt.
This route removes reference video and audio controls and requires an image before submission. The smaller input surface makes it easier to understand what will be sent to the selected model.
The page cannot submit an empty text-only task. At initialization, editing and submission boundaries it stays in frame mode and keeps only image inputs that the selected model can accept.
A generated clip must invent everything that happens after the still frame. Give the model a clean composition, a plausible motion request and one clear success criterion.
Inspect the image at the size and crop you plan to use. A clearly visible subject, readable silhouette and coherent lighting give the model a more useful starting state than a crowded collage. Leave physical room in the composition for the requested movement: a person asked to walk needs space in front of them, and a camera pullback needs surroundings that can plausibly be extended. If a product label or small facial feature must remain exact, treat the generated clip as an interpretation and review those details carefully. The image page accepts a photo, illustration or render through the visible upload controls, but it does not promise identity, typography or object geometry will remain unchanged in every frame.
The model can already inspect the uploaded frame, so use the prompt for information that is not visible: direction, pace, camera movement, environmental motion and the intended mood. A useful prompt might identify a single action, such as a cyclist beginning to move while the camera tracks sideways, and then add one or two constraints about lighting or background stability. Avoid a list of contradictory camera commands. If you want the image to remain nearly still, ask for subtle breathing, fabric, water, smoke or light movement rather than a dramatic scene change. For a larger transformation, explain the sequence in chronological order and keep it achievable within the selected duration.
Some model versions expose a second frame. Use it when the ending composition matters, such as a product arriving at a final angle or a camera move ending on another prepared view. The two frames should describe the same shot or a visually plausible transition; unrelated subjects, lighting and perspective force the model to reconcile competing instructions. Adding a last frame does not create a guaranteed frame-by-frame interpolation, and versions without end-frame support show only the opening-image slot. If you switch to such a model, the page normalizes the draft to its legal image count so a hidden second image cannot be submitted.
Preview the full result before changing the draft. Decide whether the main problem is subject motion, camera motion, composition, timing or visual stability. Change one of those factors and regenerate while keeping the image and remaining settings constant. The model selector may offer short durations from roughly five to fifteen seconds, depending on the version, and each duration asks the model to sustain a different amount of motion. Aspect ratio also changes available space around the subject. Keep successful attempts in Recent Videos so comparisons are based on actual outputs rather than memory, and check the displayed credit estimate before every new task.
Upload the opening frame, choose a compatible model and review its exact settings and credit cost before creating the video.
Required opening image
Model-specific frame controls
Recent videos and regeneration
These answers describe the current frame-based workflow and its practical limits.
Image-to-video generation creates a short clip from a still opening frame and a text prompt. The image establishes the initial subject and composition, while the prompt describes motion, camera behavior or environmental change. The selected model determines the settings and whether a final frame is available.
Yes. This dedicated route requires at least one image and stays in frame mode. If you want to generate without media, use Text to Video. If you need image, video or audio references together, use Reference to Video or the full AI Video Generator.
A final-frame control appears only for model versions configured to accept it. The opening frame is still required. Choose a final image that represents a plausible destination for the same shot; the feature does not guarantee an exact interpolation between every visual detail.
Describe the main action, its direction and pace, the camera movement and any detail that should remain stable. Focus on one shot. The image already supplies appearance and composition, so avoid spending the entire prompt repeating visible objects unless a detail is especially important.
Use a clear source with enough resolution to inspect its important details. Choose the output aspect ratio from the options shown for the selected model and crop the source with that framing in mind. The page does not claim one universal input dimension for every provider.
Credits are calculated from the selected model, resolution and duration and shown before submission. Check the live amount in the form because different versions and settings can have different costs. Available subscriptions and credit packs remain in the pricing flow.
A video model interprets and extends the still image; it does not simply move fixed pixels. Simplify the requested action, use a cleaner source, reduce conflicting details and clarify what should stay stable. Review small text, faces, hands and product geometry before publishing.
Image-driven tasks can be loaded into this form. All recent tasks remain visible, but editing a text-only or reference-media task opens the full AI Video Generator so its original inputs are preserved. Regeneration creates a new task and keeps the prior result.