Match cuts
When two shots share a shape, movement, or composition, those similarities can guide the generated bridge. Review the middle and both joins; the match is not guaranteed.
An example of what this page produces.
AI Video Transition Generator
AI video transition generator
Upload two images, describe the connecting motion, and generate a new AI video transition between them. Review the in-between frames for drift or artifacts before editing.
Pin the opening frame; add a closing frame when this workflow requires one - the model fills the motion between them.
Model
New in 2.5: generate up to 30 seconds, guide motion with as many as 30 images, or start from audio alone.
Pin the opening frame; add a closing frame when this workflow requires one - the model fills the motion between them.
The output aspect ratio follows the first-frame image automatically. To keep this image as the opening frame, crop it or extend its canvas to the ratio you need before uploading.
Using the image only as a visual reference? Choose a ratio in the AI Reference-to-Video GeneratorFree credits are for a quick preview, not the final video. Use a short duration and lower resolution to check the subject, motion, and prompt. A fuller, more compelling final shot often needs a longer duration and higher supported resolution—use paid credits when your scene needs those settings.
Audio
With audio
Synchronized sound, voice and music
Your generated video appears here. Describe a shot and hit Generate.
An AI video transition generator uses a required start image, a required end image, and a motion prompt to create new in-between video. It does not add a preset fade or wipe; review the generated middle for artifacts.
Use the last frame of the outgoing shot and the first frame of the incoming shot. Related subjects, shapes, colors, positions, or movement directions give the model a clearer path. The in-between motion is newly generated and can drift, especially when the endpoints are very different.
Two endpoint images and a prompt guide a newly generated connecting clip; inspect the result before placing it in an edit.
When two shots share a shape, movement, or composition, those similarities can guide the generated bridge. Review the middle and both joins; the match is not guaranteed.
One object can appear to become another—a sketch becoming a product, a season changing, or a face aging. When the transformation carries the message, generated motion can be more useful than a simple dissolve.
A generated bridge can suggest a camera moving from one location into another without cutting. This can support a one-take-style edit, but separate generations may still show seams, composition changes, or motion drift.
The two endpoints do most of the work. Frames that share a dominant shape, color, movement direction, or compositional anchor give the model a shorter visual journey. Unrelated images require more of the middle to be invented, which raises the chance of an unwanted morph or artifact. Choosing endpoints that rhyme is a directing decision made before generation, and it often matters more than adding adjectives to the prompt.
Duration is another useful lever. A shorter transition gives the provider less time to drift away from the endpoint guidance. Describe the intended movement instead of asking only for a transition, then review the generated middle and audio because neither a specific path nor a matching sound cue is guaranteed.
A shared shape, color, subject position, or movement direction reduces the amount the model must invent. Make this decision before generating, then verify the middle and both joins.
Begin near the shortest available duration. Add time only when the intended action feels rushed; every extra second creates more generated motion to inspect.
Model specifications
Published model limits, not marketing rounding. Every number below is enforced by the generator before your credits are reserved.
The model is ByteDance's. What FrameAI adds is what makes it usable at work: honest limits, honest billing, and output you are allowed to ship.
Text, up to 30 reference images, 10 reference videos, and 10 audio references can guide one Seedance 2.5 generation instead of being stitched together afterwards.
When audio is enabled, Seedance 2.5 can generate footsteps, room tone, impacts, dialogue, and score with the video. Review timing and content before publishing, or switch audio off when the clip needs separate sound design.
One prompt covers a 9:16 vertical cut for Reels, a 16:9 master for YouTube, and a 21:9 anamorphic frame for a title sequence. Resolution runs from 480p to 4K, so the same take holds up on a cinema screen and on a phone.
The maximum credit reservation is calculated from the selected model, duration, resolution, ratio, and video-reference input, then shown before submission. A successful video job can settle lower from provider-reported usage; confirmed failed jobs return the reservation.
Use Mini for cost-efficient prompt exploration, Fast when you want a balance of generation speed and cost, and Seedance 2.5 when its longer duration or richer reference inputs fit the shot. Switching models starts a new generation, so results may vary.
Content generated with a paid subscription or purchased credit pack may be used commercially under the Terms of Service. Content generated solely with free sign-up or promotional credits is for evaluation and personal use only.
Two endpoint frames in, one candidate bridge to review.
Use the last frame of the outgoing shot and the first frame of the incoming one. Look for something they share; that element can guide the generated transition.
Say how the camera or the subject gets from one to the other. "Pushes through", "rotates into", "dissolves as the shape holds" all give the model a path to follow.
Start near the shorter end of the available duration range, then place the result between the two shots in your edit. If it wanders, shorten it before rewriting the prompt.
Each of these needs the audience to feel a connection that a straight cut would break.
Before and after, sketch to finished object, closed to open. The change is the message, so showing the change beats showing two states.
Continuous movement through several environments is common in title design. Generating the links can avoid hand-animating every intermediate frame, but the new motion may drift and must be reviewed.
When a story spans several generated shots, a transition clip can bridge one chosen endpoint to the next. Review both joins and keep only the connections that support the edit.
It is a new generated clip that uses a required start frame, a required end frame, and your transition prompt to propose the movement between two shots. The middle is generated and should be reviewed.
A crossfade blends two images by opacity. This workflow asks the provider to generate new intermediate frames from the two endpoints and the prompt, so the result may show movement rather than a simple dissolve, but it can also introduce artifacts.
The default model on this page supports 4 to 30 seconds. Start near the shorter end for a simple connection, then add time only if the change feels rushed. Longer clips create more generated middle to inspect.
The endpoints may be too different, or the provider may interpret the prompt in an unexpected way. Try frames that share a shape, color, subject position, or movement direction; shorten the duration; and make the transition prompt more specific.
Yes. Upload them as endpoint images; they can come from FrameAI generations or footage you filmed.
They share a mechanism but not a purpose. First & Last Frames is for authoring one shot whose start and end are guided by images. Video Transition is for joining two shots that already exist.
Each tool uses Seedance 2.5 for a different job. Move between them freely — your credits, history, and exports are shared.
Choose paid access when you need commercial-use eligibility, predictable generation costs, and credits restored after a failed or rejected job. One-time packs are available with no subscription.