Match cuts
Two shots sharing a shape, a movement or a composition, joined so the eye follows the similarity across the cut. Generating the in-between makes the match land instead of merely implying it.
Video Transition
AI video transitions
Provide a start frame and end frame with a transition prompt — FrameAI generates a smooth AI video cut between two shots. Review the generated bridge before editing.
Pin the opening frame; add a closing frame when this workflow requires one — the model fills the motion between them.
Model
Pin the opening frame; add a closing frame when this workflow requires one — the model fills the motion between them.
0/2000
Outputs
Audio
With audio
Synchronized sound, voice and music
Your generated video appears here. Describe a shot and hit Generate.
Provide a start frame, an end frame, and a transition prompt. The provider generates a new 4-to-15-second clip intended to connect the endpoints. It is not a crossfade and it is not a guarantee that every intermediate frame will look natural, so review the result before editing it in.
Use endpoints with a related subject, shape, colour, or movement when possible, then describe the action or camera path you want to see between them. The two frames are required inputs; the in-between motion is newly generated and can drift, especially when the endpoints are very different.
If a straight cut works, use a straight cut. These are the cases where it does not.
Two shots sharing a shape, a movement or a composition, joined so the eye follows the similarity across the cut. Generating the in-between makes the match land instead of merely implying it.
One object becoming another — a sketch becoming a product, a season changing, a face ageing. The transformation itself carries the message, which is why it cannot be a dissolve.
A camera that appears to move from one location into another without cutting. The oner effect, assembled from separate generations rather than shot in one impossible take.
The two ends do most of the work. A transition between frames that already share something — a dominant shape, a colour, a direction of movement, a compositional anchor — is a short journey, and the model executes it convincingly. A transition between two unrelated images is a long journey, and it will either look like a morph for its own sake or fall apart in the middle. Choosing endpoints that rhyme is a directing decision made before you generate anything, and it matters more than the prompt.
Duration is the second lever, and shorter is almost always better. A transition is connective tissue; the audience should feel it rather than study it. Four or five seconds is usually plenty, and stretching one out gives the model more room to wander away from both endpoints. Describe the nature of the movement — "the camera pushes through the window and the interior resolves" — rather than just asking for a transition, and remember that the generated audio will follow the movement, so a fast transition brings its own whoosh.
A shared shape, colour or movement direction turns a difficult transition into an easy one. This decision is made before generating, not fixed afterwards.
Four to five seconds reads as intentional. Longer transitions invite the model to drift away from both frames you pinned.
Model specifications
Published limits, not marketing rounding. Every number below is enforced by the generator before your credits are reserved, so a job that would exceed a limit is rejected instead of failing halfway.
The model is ByteDance’s. What FrameAI adds is the part that decides whether it is usable at work: honest limits, honest billing, and output you are allowed to ship.
Text, up to nine reference images and up to three reference video clips are fused in one generation, not stitched afterwards. There is no separate text-to-video mode to switch into — you add whatever material you have and the model reconciles it.
Seedance 2.0 writes footsteps, room tone, impacts and score alongside the frames, locked to the action. Most models hand you a silent clip and leave the sound design to you; here the export is already finished.
One prompt covers a 9:16 vertical cut for Reels, a 16:9 master for YouTube and a 21:9 anamorphic frame for a title sequence. Resolution runs from 480p up to 4K, so the same take can serve a cinema-shaped frame or a phone.
The exact credit cost is calculated from your duration, resolution and ratio and shown before you press generate — never after. If a generation fails on the provider side, the reserved credits are returned automatically.
Seedance 2.0 Fast and Seedance 2.0 Mini exist for the twenty drafts nobody sees. Rough a shot at 480p for a fraction of the credits, then re-run the take that worked at full resolution on the flagship model.
Paid-plan output may be used commercially under the Terms of Service. Free-credit output is for evaluation only.
Two frames in, one piece of connective movement out.
The last frame of the outgoing shot and the first frame of the incoming one. Look for something they share — that shared element is what the transition will travel along.
Say how the camera or the subject gets from one to the other. "Pushes through", "rotates into", "dissolves as the shape holds" all give the model a path to follow.
Keep it to four or five seconds, then drop it between the two shots in your edit. If it wanders, shorten it before rewriting the prompt.
Each of these needs the audience to feel a connection that a straight cut would break.
Before and after, sketch to finished object, closed to open. The change is the message, so showing the change beats showing two states.
Continuous movement through several environments is the house style of title design, and generating the links is dramatically cheaper than animating them.
When a story needs more than fifteen seconds, generated transitions turn a set of separate takes into something that reads as continuous.
It is a new generated clip that uses a required start frame, a required end frame, and your transition prompt to propose the movement between two shots. The middle is generated and should be reviewed.
A crossfade blends two images by opacity. This workflow asks the provider to generate new intermediate frames from the two endpoints and the prompt, so the result may show movement rather than a simple dissolve, but it can also introduce artifacts.
Four to five seconds in most cases. Transitions are connective tissue, and a longer one both draws attention to itself and gives the model more room to drift away from your two endpoints.
The endpoints may be too different, or the provider may interpret the prompt in an unexpected way. Try frames that share a shape, colour, or movement direction, shorten the duration, and make the transition prompt more specific.
Yes. Upload them as reference images — the transition does not care whether the endpoints were generated here or filmed on a camera.
They share a mechanism but not a purpose. First & Last Frames is for authoring a single shot whose start and end you want to control. Video transition is for joining two shots that already exist.
Each tool is the same Seedance 2.0 model pointed at a different job. Move between them freely — your credits, history and exports are shared.
Write one line, pick a ratio, press generate. No editing suite, no render farm, no post-production pass — a finished clip with sound, ready to download.