01
What separates a prompt that works from one that does not
Concrete visual detail beats adjectives, every time. “A beautiful beach scene” gives a model nothing to decide; “slow dolly-in at golden hour, low camera, wet sand reflecting the sky, one figure walking away from the lens” gives it four decisions already made.
So each prompt this returns names four things explicitly: the shot and camera movement, what the subject is doing, the lighting, and the mood. That is the structure most text-to-video models respond to, and writing it out is most of the craft.