Upload one photo, describe the motion, and get a clip back — with audio on most models. Thirteen engines behind one box, free drafts to find out which one suits your shot.
Input
JPG · PNG · WebP · HEIC, 20 MB
Models available
13, selectable per job
Clip length
5–30 s depending on model
Audio
Native on most models
Output
Up to 1080p+ on credits
Free tier
3 drafts daily, no account
Typical time
Seconds to 3 minutes
Our one-line verdict: the highest-hit-rate way to use AI video. A photo hands the model composition, lighting and subject for free — image-to-video succeeds where text-to-video is still guessing.
Animating a still is the most reliable thing AI video does, and the reason is structural: your photo already settles composition, lighting, subject identity and framing, so the model only has to solve motion. That's a far smaller problem than inventing a whole scene from a sentence. The catch is that models differ enormously here — one treats your image as the literal first frame, another as a loose suggestion — so this page is mostly about routing you to the right one. Free drafts exist so you can find out which without paying for the lesson.
The honest routing table. Every one of these is selectable in the generator above — start on the free tier, then take the link when you know what you need.
Its architecture treats your still as literally frame one, so composition, light and identity survive. Took the image-to-video arena crown on exactly this, and renders in seconds.
Physical realism plus ambience and dialogue in the same pass. Eight seconds, but the eight seconds hold up when a client leans in.
Thirty seconds in one take, with a supplied soundtrack able to drive the pacing. When the animation needs to breathe rather than blink.
Reference-locked identity: feed it up to seven images and the person, product or logo stays itself shot after shot. Series work lives here.
Pin up to sixteen keyframes and choreograph exactly where each moment lands — plus HDR/EXR output a colorist can grade.
Draft free on Seedance Fast, then step up to 2.0 for shippable clips with native audio at the family's lowest paid rate.

JPG, PNG, WebP or HEIC up to 20 MB. Sharp, well-lit images with a clear subject animate best — a blurry source stays blurry once it moves.

The photo already carries the scene. Spend the prompt on movement: what the camera does, what the subject does, what we hear. 'Slow push-in, hair moves in the breeze, distant traffic' beats re-describing the picture.

Start on the free draft tier to check the motion reads right, then re-run the same prompt on the model the table below points you to. Download when it lands.
The right column is the part most tools leave out. Knowing the failure modes before you upload saves more time than any prompt guide.
• Motion beats invention.
Because the photo fixes composition and lighting, the first attempt is usually usable. Hit rates here are dramatically higher than text-to-video.
• Audio in the same pass.
Most models on this page generate ambience, effects or dialogue alongside the picture, so an animated still arrives publish-ready rather than silent.
• Your subject, not a lookalike.
Reference-driven models keep the actual product or person from your photo — the difference between animating your asset and generating a similar one.
• Cheap to try.
Three free drafts a day, no account. Motion ideas are quick to test and cheap to abandon, which is the correct way to work.
• Faces and hands under big motion.
Ask for dramatic movement and faces can drift or hands can misbehave. Subtle motion prompts are not a limitation dodge — they genuinely produce better results.
• Low-resolution or blurry sources.
Animation doesn't add detail that isn't there. A soft source produces a soft clip; start from the sharpest version you have.
• Complex scene changes.
A still can be animated, not rewritten. If you need a different setting, generate the video from text instead — that's the next page over.
• Photos of people who didn't agree.
Same rule as everywhere on this site: your own photos or ones you have the right to use, no public figures, no minors.
One packshot becomes a scrolling-stopping clip: slow orbit, light sweep, a hand entering frame. Cheaper than a video shoot and reuses assets you already paid for.
Subtle movement on a family portrait — breath, a shift of weight, drifting light. Restraint is everything here; the prompt recipes on the model pages lean that way for a reason.
Feeds reward motion. Animating an existing hero image is the fastest route from a static campaign to a video-native one without a reshoot.
Upload the image, write a short prompt describing the motion rather than the scene, pick a model, and generate. The whole loop takes under a minute on the free draft tier, and the model-routing table above tells you which engine suits your shot.
No photo to start from? Describe the whole scene instead.
Already have footage — restyle or edit it.
Every engine, with one line each on what it wins.