Push and pull
Dolly-in, push-in, pull-back, zoom. Moving toward a subject concentrates attention and builds tension; moving away reveals context and releases it. The most useful pair in the vocabulary, and the easiest to overdo.
An example of what this page produces.
Motion Control
AI camera motion control for video
Direct camera moves in AI video — dolly-in, orbit, crane-up, whip pan, handheld shake — using natural language prompts. Free AI camera motion control, no trajectory software needed.
MP4 or MOV, 2-15s each, 15s total, up to 200MB. Requires R2 with a public domain - the model fetches the file over the network.
Model
New in 2.5: generate up to 30 seconds, guide motion with as many as 30 images, or start from audio alone.
Add at least one reference below. Images, videos, and audio can be used alone or combined.
Reference videos (optional) 0/10
MP4 or MOV, 2-30s each, 30s combined, up to 200MB. Requires R2 with a public domain - the model fetches the file over the network.
Reference audio (optional) 0/10
Seedance 2.5 accepts MP3 or WAV clips, 2-30s each and 30s combined. Audio-only reference is supported.
Free credits are for a quick preview, not the final video. Use a short duration and lower resolution to check the subject, motion, and prompt. A fuller, more compelling final shot often needs a longer duration and higher supported resolution—use paid credits when your scene needs those settings.
Audio
With audio
Synchronized sound, voice and music
Your generated video appears here. Describe a shot and hit Generate.
Upload a source clip and a prompt, then specify the camera motion — dolly, pan, orbit, crane, or handheld. The generator follows your direction to produce a new clip, without trajectory software.
Camera language can make a prompt more specific: state the move, its speed, the subject action, and where the shot should begin or end. The generated result may still choose a different path or timing, so treat the prompt as direction rather than a tracking or motion-control file. Review the clip and refine the wording when the move drifts.
Natural language commands for camera motion, from locked-off shots to complex crane moves.
Dolly-in, push-in, pull-back, zoom. Moving toward a subject concentrates attention and builds tension; moving away reveals context and releases it. The most useful pair in the vocabulary, and the easiest to overdo.
Arc, orbit, whip pan, tilt. Lateral movement describes an object in three dimensions - which is why almost every product film in existence orbits. A whip pan does the opposite job: it hides a cut.
Crane up for scale, handheld for immediacy and unease, locked-off for formality, rack focus to move attention within a static frame. These set the register of a shot before anything happens in it.
Name one move per shot. A prompt that asks to dolly in, orbit, crane up and rack focus in eight seconds is asking for chaos, and chaos is what comes back. Pick the single move that does the storytelling job and describe its speed as well as its type - "slow push-in" and "fast push-in" are different shots with different emotional readings, and speed is the qualifier people most often leave out.
Give the requested move a subject and a direction. Pair camera language with an action and, where useful, a starting and ending relationship. Generated motion, timing, and audio can still differ from the prompt, so review the result rather than expecting a fixed path or sound cue.
"Slow orbit, left to right" is directable. "Dynamic cinematic camera movement" is not - it names no move and no speed, so the model chooses both.
Stating where the shot starts and where it ends up produces far more controlled results than naming the move alone.
Model specifications
Published model limits, not marketing rounding. Every number below is enforced by the generator before your credits are reserved.
The model is ByteDance's. What FrameAI adds is what makes it usable at work: honest limits, honest billing, and output you are allowed to ship.
Text, up to 30 reference images, 10 reference videos, and 10 audio references can guide one Seedance 2.5 generation instead of being stitched together afterwards.
When audio is enabled, Seedance 2.5 can generate footsteps, room tone, impacts, dialogue, and score with the video. Review timing and content before publishing, or switch audio off when the clip needs separate sound design.
One prompt covers a 9:16 vertical cut for Reels, a 16:9 master for YouTube, and a 21:9 anamorphic frame for a title sequence. Resolution runs from 480p to 4K, so the same take holds up on a cinema screen and on a phone.
The maximum credit reservation is calculated from the selected model, duration, resolution, ratio, and video-reference input, then shown before submission. A successful video job can settle lower from provider-reported usage; confirmed failed jobs return the reservation.
Use Mini for cost-efficient prompt exploration, Fast when you want a balance of generation speed and cost, and Seedance 2.5 when its longer duration or richer reference inputs fit the shot. Switching models starts a new generation, so results may vary.
Content generated with a paid subscription or purchased credit pack may be used commercially under the Terms of Service. Content generated solely with free sign-up or promotional credits is for evaluation and personal use only.
Same generator, one deliberate addition to the prompt.
Ask what the camera should make the viewer feel - pressure, scale, urgency, formality - and pick the move that does that job. One per shot.
Slow or fast, and from where to where. Then describe the subject action the move is built around, so the camera has something to be about.
If the move reads wrong, adjust the speed qualifier before rewriting anything else - it is usually the variable at fault.
In each of these the move is not a flourish - it is the reason the shot works.
A controlled orbit or push-in is the standard grammar of showing an object properly. Random drift makes the same product look incidental rather than considered.
Crane and pull-back moves are how landscapes, buildings and crowds read as large. Without the move, scale simply does not register.
A slow push-in during a held moment does what no amount of description can. Camera speed is the most direct control you have over how a clip feels.
It is a prompt-based workflow for describing camera movement. You name the move and its speed, optionally provide a source video, and the provider generates a new clip using that information as guidance.
You can describe common camera language such as dolly or push-in, pull-back, orbit, pan, whip pan, tilt, crane, handheld, locked-off, and rack focus. These are prompt cues, not a formal motion-path input.
You can, but it rarely helps. A four-to-fifteen-second clip has room for one move executed well. If a sequence needs several, generate them as separate shots and join them with Video Transition.
Say it - "slow push-in", "fast whip pan", "gentle orbit". Speed is the qualifier most often omitted and the one that most changes how a shot reads, so it is worth stating on every camera instruction.
Yes. This page requires a reference video, which gives the provider source motion and visual context; the camera words in your prompt add guidance. Neither input guarantees that the subject or path will remain identical.
The provider generates a new clip and may interpret the wording differently. Make the prompt concrete - for example, "slow dolly-in from wide to medium" - then regenerate and compare when the first result drifts.
Each tool uses Seedance 2.5 for a different job. Move between them freely — your credits, history, and exports are shared.
Choose paid access when you need commercial-use eligibility, predictable generation costs, and credits restored after a failed or rejected job. One-time packs are available with no subscription.