TL;DR: A Redditor tested a stack of video prompts across PixVerse, Seedance 2.5, and MiniMax H3, then kept the five that actually held together. One continuous camera move, age changes, a woman turning into a cat, an FPV flight into a dental clinic, and a runaway businessman. Copy the structure, swap the details.
u/IcyPea3192 posted the breakdown on r/PromptEngineering, and it’s worth a slow read if you’ve ever fed a video model a vague idea and gotten mush back. The original poster didn’t just paste prompts and walk away either. Each one is built the same way: one continuous shot, a strict timeline in seconds, and a locked description of what stays the same while everything else changes around it.
That structure is the actual lesson here, more than any single prompt.
Why These Hold Together
Video models drift. Give them fifteen seconds and a loose idea, and by second eight the face has changed, the lighting has reset, or the camera has teleported somewhere new. This Redditor’s fix is to lock three things in every prompt: the camera behavior, the identity anchor, and the second-by-second timeline.
Look at the age transition prompt. It doesn’t say “show someone aging.” It says one slow 360-degree orbit, constant speed, same face and body proportions preserved, then a hard cut of ages tied to exact seconds with a line of dialogue at each stage. The model has almost no room to improvise the wrong thing.
The cat transformation does the same trick with an object instead of a face: the black dress stays visible on the ground the whole time, so the model has a physical anchor to track through the smoke effect.
The Five Prompts
1. Age progression, one orbit, one person
Style is cinematic realism with restrained emotion, natural skin texture and a golden-hour-to-night lighting transition. One continuous shot with no cuts.
The camera performs one slow, steady clockwise 360-degree orbit around the same person standing in an open field. The face, eyes, core identity and body proportions remain recognizably the same while age changes progressively. The orbit speed remains constant.
0 to 3 seconds. Age 8. A child stands in tall grass holding a paper airplane in warm golden sunlight. The child smiles and says, “I’m going farther than anyone.”
3 to 6 seconds. Age 16. The same person becomes a teenager carrying a school backpack. The sunset deepens. Hair and clothing change naturally during the orbit. The teenager says, “I’m not giving up.”
6 to 9 seconds. Age 25. The same person becomes a young adult in a simple jacket. The sky turns blue at dusk. The adult looks uncertain and says, “I hope I chose right.”
9 to 12 seconds. Age 45. Subtle wrinkles and grey hair appear as night begins settling over the field. The same person breathes deeply and says, “I did what I could.”
12 to 15 seconds. Age 75. The same person is elderly beneath a clear starry sky. The orbit finishes in front. The elderly person looks upward, then back toward the camera and says, “And it was enough.” Hold the final expression for one second.
Audio includes gentle wind in tall grass, distant birds fading into night insects and restrained piano entering at 9 seconds. No subtitles, logo or text overlays.
2. Woman into black cat
Create a fifteen-second cinematic transformation video in one continuous shot with realistic live action, blue-hour city lighting, wet reflective pavement, light rain, no cuts and no visible injuries.
0 to 3 seconds. A woman in a flowing black dress walks toward the camera on a quiet rain-damp street. The camera tracks backward at waist height. She stops beneath a streetlamp and looks over her shoulder.
3 to 6 seconds. Black smoke rises from the pavement and wraps around her body. The camera tilts downward. She disappears inside the smoke, leaving the empty black dress suspended for half a second before it collapses onto the wet ground.
6 to 10 seconds. The camera pushes closer to the fallen dress. The fabric shifts. Two amber eyes appear between the folds. A large black cat emerges from beneath the dress, steps onto the pavement and looks into the lens. Keep the dress visible behind it.
10 to 13 seconds. The cat turns and walks down the reflective street. The camera drops to a low tracking angle and follows beside it. Keep the black fur, amber eyes and body size consistent.
13 to 15 seconds. The cat stops at a narrow dark alley, looks back for one second and runs into the darkness. A faint swirl of black smoke appears inside the alley. Hold on the empty entrance.
Audio includes rain, distant traffic, paw steps and fabric rustle, with one low sound when the cat looks back. No music until the last two seconds.
3. FPV flight, skyline to dental clinic
Use Image 1 as the exact starting frame and Image 2 as the exact final frame. Create one continuous first-person FPV camera shot lasting fifteen seconds. Never show a drone, propellers, camera shadow or physical rig. Use fast but controlled forward motion, gentle curves, subtle banking and a stable horizon. No flips, full rotation, teleporting or cuts.
0 to 3 seconds. Begin high above the Manhattan skyline. Accelerate forward, skim past two rooftops and descend toward a glass office tower.
3 to 6 seconds. Pass through an open window into the office. Fly between desks, monitors and chairs, curve around one central column and exit through an open window on the opposite side.
6 to 10 seconds. Descend along the exterior to street level. Glide above traffic, passing yellow taxis, a city bus and pedestrians without collisions. Follow the road toward a Times Square-like commercial area.
10 to 13 seconds. Rise toward illuminated billboards, make one curved turn around a screen and descend toward street level. Reveal the dental clinic matching Image 2.
13 to 15 seconds. Pass through the clinic entrance, glide between the reception desk and waiting chairs, slow during the final half-second and settle into the framing of Image 2.
Audio moves from wind and traffic to office room tone, street ambience and a quiet clinic hum. No dialogue or music.
4. Performers breaking into laughter
Show an original pop-group performance under purple and pink LED lights. Three adult East Asian women with distinct faces and consistent outfits stand in a line wearing headset microphones. The left performer wears a lime-green sleeveless top, the center performer wears a baby-pink crop top with white piping, and the right performer wears a navy-and-green striped collared top. Do not resemble real celebrities.
0 to 3 seconds. They dance lightly in sync. The center performer flips her hair. The right performer notices something funny off-camera but keeps performing.
3 to 7 seconds. The right performer starts laughing. The other two notice and all three bring both hands to their mouths at slightly different moments while their shoulders shake. Their feet keep the choreography going.
7 to 12 seconds. They recover but continue giggling, exchanging side glances while staying in rhythm. Hold a medium handheld concert framing that shows all three.
12 to 15 seconds. On the beat, all three change from covering their mouths to double peace signs near their faces. They look at the camera and hold the final pose.
Use an original pop instrumental, crowd cheer and brief headset-mic laughter. No recognizable song, brand logos or text overlays.
5. Businessman versus the ticket
Create a fast-paced street comedy in realistic live action. An adult businessman in a tailored bright-blue suit realizes that a city officer is writing him a small jaywalking ticket.
0 to 3 seconds. The businessman notices the ticket pad, looks horrified, checks both directions and takes two steps backward while maintaining eye contact. The officer raises one eyebrow.
3 to 7 seconds. The businessman turns and sprints down the sidewalk holding his briefcase. The officer hesitates, sighs and jogs after him. Pedestrians move aside while he runs with action-hero intensity.
7 to 11 seconds. He reaches a tiny puddle. Time shifts into slow motion. He leaps over it as if escaping an explosion. An orange reflection flashes in the puddle, his briefcase opens and harmless office papers scatter around him. Return to normal speed as he lands.
11 to 13 seconds. He turns a corner, celebrates briefly and straightens his tie while jogging. The camera moves in front of him and reveals a second officer waiting with another ticket pad.
13 to 15 seconds. He skids to a stop, raises both hands and looks into the camera. Both officers arrive on either side. Hold the final expression.
Audio includes footsteps, the briefcase latch, papers, one dramatic boom over the puddle and silence at the final stare. No dialogue, brand logos or injuries.
One commenter zeroed in on why the age progression works: the dialogue at each stage isn’t decoration, it’s what ties the visual change to an actual emotional arc instead of just a face filter running on a timer.
Use Cases 🎬
- Brand storytelling: the age-progression structure works for any “journey” ad, product loyalty over decades, a founder story, a before-and-after transformation.
- Horror or fantasy shorts: the cat transformation shows how to hide a hard cut inside smoke or fabric so the model has less to fake.
- City or venue tours: the FPV prompt is a template for any “fly through our building” or “tour our city” clip, just swap the landmarks and the final destination image.
- Comedy skits: the officer chase proves you can chain three separate beats (fear, chase, slapstick) into one shot as long as the seconds are locked down.
Prompt of the Day
If you only steal one structure, take the age-progression skeleton and reuse it for anything with a life-stage arc: a company’s ten-year growth, a product’s evolution across versions, a relationship over decades. Lock the camera move, lock the identity anchor, assign one line of dialogue per stage, and let the model fill in the visuals.
Go read the full thread and the comments on r/PromptEngineering, then try the age-progression prompt with your own subject and see how close your model gets on the first pass.
Five video prompts that gave me results worth keeping
by u/IcyPea3192 in PromptEngineering