How do you generate AI video from a photo?
Upload the photo to an image-to-video tool, write a short prompt describing only the motion you want (not the scene — the scene is your photo), and generate. The model uses your photo as the first frame and animates from there. Best image-to-video models in 2026: Veo 3, Kling 2 and Runway Gen-3.
- Prompt motion only, not the scene.
- Veo 3 / Kling 2 / Runway Gen-3 lead image-to-video in 2026.
- Higher-resolution input photos give better results.
- Square-on photos work better than extreme angles.
The prompting trick most beginners miss
With text-to-video you describe the whole scene. With image-to-video, the scene is already locked by your photo — the model just needs to know what should move. Prompt only motion: 'slow camera push-in, subject smiles, wind blows hair gently'. Re-describing the scene confuses the model.
Photo quality matters
The output cannot be higher quality than the input. Use a clean, well-lit photo of at least 1080×1920 for vertical or 1920×1080 for horizontal. Avoid heavy filters or compressed Instagram screenshots — the model amplifies the artifacts.
FAQ
Can I use any photo I want?
Photos of yourself, your products, or licensed stock = yes. Photos of other real people or copyrighted characters = legally risky and against most platform terms.