Guide the ending with an image
A final-frame reference gives the model a clear target for the ending. It guides the result but does not guarantee a pixel-perfect arrival, so review the final frames.
An example of what this page produces.
First & Last Frame Video Generator
First & last frame AI video generator
Upload a required first frame and last frame, describe the motion, and generate a new AI video clip using both images as endpoint guides. Free first and last frame AI video generator.
Pin the opening frame; add a closing frame when this workflow requires one - the model fills the motion between them.
Model
New in 2.5: generate up to 30 seconds, guide motion with as many as 30 images, or start from audio alone.
Pin the opening frame; add a closing frame when this workflow requires one - the model fills the motion between them.
The output aspect ratio follows the first-frame image automatically. To keep this image as the opening frame, crop it or extend its canvas to the ratio you need before uploading.
Using the image only as a visual reference? Choose a ratio in the AI Reference-to-Video GeneratorFree credits are for a quick preview, not the final video. Use a short duration and lower resolution to check the subject, motion, and prompt. A fuller, more compelling final shot often needs a longer duration and higher supported resolution—use paid credits when your scene needs those settings.
Audio
With audio
Synchronized sound, voice and music
Your generated video appears here. Describe a shot and hit Generate.
A first and last frame to video generator uses one required opening image and one required ending image as endpoint guidance, then generates the motion between them from your prompt. The middle and final pixels can drift, so review the full clip.
Use this workflow to author one shot with a planned beginning and ending: a product reveal, a pose change, or a loop concept. If you are joining the last frame of one existing shot to the first frame of another, use Video Transition instead. In either workflow, the supplied frames are guidance rather than a pixel-perfect guarantee.
Two endpoint images and a motion prompt guide a newly generated connecting clip; the result still needs review.
A final-frame reference gives the model a clear target for the ending. It guides the result but does not guarantee a pixel-perfect arrival, so review the final frames.
Use the same image as both the first and last frame to guide the clip back toward its starting composition. Generated frames may drift, and the join may still need trimming.
Related endpoint images can guide a connected transition between shots. Review the generated ending and opening because the join may still need trimming.
The distance between your two frames sets the difficulty. Images with matching orientation, similar subject scale, viewpoint, and lighting describe a shorter visual journey. Large changes force the model to invent more of the middle. If the endpoints are far apart, simplify the action, allow more time, or split the idea into separate shots.
Duration is the other half of the equation. Short settings leave less generated middle to inspect but can rush a large change; long settings give complex motion more room but also create more opportunities for drift. Match duration to the visual distance, describe the camera and subject path, and crop both images to the intended output orientation before generating.
Shared subject, similar viewpoint, consistent lighting, and matching orientation reduce the amount the model must invent.
Start near the shortest available duration. Add time only when the intended camera or subject movement feels rushed, and inspect the entire middle after every change.
Model specifications
Published model limits, not marketing rounding. Every number below is enforced by the generator before your credits are reserved.
The model is ByteDance's. What FrameAI adds is what makes it usable at work: honest limits, honest billing, and output you are allowed to ship.
Text, up to 30 reference images, 10 reference videos, and 10 audio references can guide one Seedance 2.5 generation instead of being stitched together afterwards.
When audio is enabled, Seedance 2.5 can generate footsteps, room tone, impacts, dialogue, and score with the video. Review timing and content before publishing, or switch audio off when the clip needs separate sound design.
One prompt covers a 9:16 vertical cut for Reels, a 16:9 master for YouTube, and a 21:9 anamorphic frame for a title sequence. Resolution runs from 480p to 4K, so the same take holds up on a cinema screen and on a phone.
The maximum credit reservation is calculated from the selected model, duration, resolution, ratio, and video-reference input, then shown before submission. A successful video job can settle lower from provider-reported usage; confirmed failed jobs return the reservation.
Use Mini for cost-efficient prompt exploration, Fast when you want a balance of generation speed and cost, and Seedance 2.5 when its longer duration or richer reference inputs fit the shot. Switching models starts a new generation, so results may vary.
Content generated with a paid subscription or purchased credit pack may be used commercially under the Terms of Service. Content generated solely with free sign-up or promotional credits is for evaluation and personal use only.
Two images and a sentence about the path between them.
The composition the shot starts on. Sharp, well-lit and framed with a little room for the camera to move.
The ending the edit requires. For a loop concept, you can reuse the opening image, then review the generated motion and actual join.
Say how the shot travels between the two—the camera move and what the subject does—and set a duration proportional to the visual distance.
Each of these has a requirement that a text prompt alone cannot guarantee.
A supplied closing frame gives the provider explicit endpoint guidance, but the generated output can still differ at the final pixels. Review the ending against brand requirements.
Using the same image at both ends can guide the clip toward a loop, but the generated middle and join may drift. Review and trim the result before publishing.
Related endpoint images can guide a more connected edit, but generated motion and the actual join may drift and require trimming.
You supply a required opening image and closing image, then the provider generates a new clip between them from your prompt. The two images guide the endpoints; the generated middle can still vary.
You can try using the same image at both ends to guide a loop, but the generated middle and the actual join still need review. It is not a guarantee of a perfect loop.
The model has to invent a long path between them, and that invention shows up in the middle of the clip. Either choose endpoints that share a subject, viewpoint or lighting scheme, or split the idea into two shots joined by a transition.
The default model on this page supports 4 to 30 seconds. Start near the shorter end for a small visual change, then add time only if the motion feels rushed. Longer clips create more generated middle to inspect.
Yes, and you should. The frames set the endpoints; the prompt sets the path—the camera move and subject action that get you from one to the other.
The mechanism is similar, but the intent differs. First and Last Frames uses two endpoint images as guidance for one newly generated shot. Video Transition uses endpoint images from two existing shots to guide a new connecting clip. Neither guarantees an exact endpoint match.
Each tool uses Seedance 2.5 for a different job. Move between them freely — your credits, history, and exports are shared.
Choose paid access when you need commercial-use eligibility, predictable generation costs, and credits restored after a failed or rejected job. One-time packs are available with no subscription.