How to make an AI video that looks real
Making an AI video is easy now. Making one that people don't immediately clock as AI takes a few deliberate choices, and most of them come down to asking for less: less motion, less polish, fewer things happening at once. This guide covers the choices that matter, then the usual artifacts and what causes each one.
Three ways to make an AI video
- Text to video: you describe the shot and the model invents everything. Quick for exploring ideas, unreliable for anything that has to match a real product or person.
- Image to video: you supply the first frame and the model animates it. You control the opening, but anything that isn't in that frame, such as the back of a box or a face turning, is a guess.
- Reference to video: you supply separate images of the person, the product and the place, and describe what happens. The model composes the shot itself while holding each element to its reference. For ads, this is usually the best fit.
Ask for one action
Realism falls apart under complexity. One person doing one thing in one place, with at most one camera move, looks far more convincing than several actions crammed into a few seconds. If you need two actions, give each its own stretch of time and say exactly where the first one ends. The AI video ad prompt template shows how to write those stretches.
Choose motion the model handles well
Some movements come out convincing far more often than others. Design the shot around the reliable ones.
- Usually reliable: picking something up, turning toward the camera, taking a sip, smiling, a few steps of walking, a slow push in.
- Often break: fast hand movements, fingers interlacing, liquids poured quickly, objects passing behind one another, text on something that moves.
Make it look shot on a phone
Video that looks too perfect reads as generated. Real social video has:
- a handheld camera with slight, natural sway,
- a lens about as wide as a phone camera, at arm's length or eye level,
- window light or a single lamp rather than studio lighting,
- visible skin texture, stray hairs and some everyday clutter in the background,
- an ordinary place: a kitchen, a parked car, a bathroom shelf.
All three of these are generated, and all three use those cues: a phone at arm's length, natural light, and real places with weather and crowds in them.
Write these into your style direction specifically. "Soft window light from the left, phone camera at arm's length, natural skin texture" does far more than "cinematic, 8K, ultra-detailed", which pushes the other way.
Don't forget the sound
A silent video, or a voice with no room around it, gives AI away as fast as anything visual. Ask for room sound that matches the place: traffic outside a car, the hum of a kitchen. Keep music low under speech, or leave it out. If someone talks, write their exact words and keep it to a sentence; the longer the speech, the more chances for the lips and the words to drift apart.
Common artifacts and their causes
- Garbled text on labels, signs or screens: no clear reference for the text, or the model was asked to write text itself. Give a close-up of the real label, and add captions afterwards rather than asking the model for on-screen words.
- A product that changes shape or size: the model only had one view and no sense of scale. Say how big it is next to a hand, and give more than one angle if you have them.
- A face that drifts: one portrait isn't enough to hold an identity through movement. Use a character sheet that shows the person from several angles.
- Extra fingers or strange grips: describe the grip ("held in her left hand, label facing the lens") and avoid close-ups of fiddly hand work.
- Floaty, weightless motion: too much movement for the length of the clip. Slow it down or cut an action.
- Extra people or objects: list what must not appear. Exclusions are part of the prompt, not an afterthought.
Keep it short
Keep your first ad focused on one action, even when the model supports longer clips. In fairytale video, Seedance 2.5 supports 4 to 30 seconds. A longer ad needs a timed plan with room for each action and the dialogue, rather than a longer paragraph asking for everything at once.
Check it before anyone else sees it
Watch at full size, then step through the moments where hands touch the product or the face turns: that's where problems hide. If you're making an ad, the review checklist in how to make an AI ad covers what else to look for.
Making AI videos of your product in fairytale video
fairytale video makes reference-to-video ads. Your product, an optional creator and a scene go in as references, and the studio writes a timed plan for each video before rendering. You choose 4 to 30 seconds, 9:16 or 16:9, and 720p or 1080p. Every video is a paid Seedance 2.5 render, with its credit price shown before you start, and keeps the exact prompt and reference images it was made from. Review the result before posting. If you are comparing tools, how to choose an AI ad generator lists what to test.