Script free with an account·render in about a minute·no watermark on any plan
AI Story Video Generator
A story told in five beats, illustrated frame by frame, narrated and captioned.
Start free — no credit card, and the script and the shot list cost no credits.
01Give it the premise. A situation and, if you have it, the turn. One or two lines is plenty.
02Five beats, five frames. Situation, complication, escalation, turn, resolution — each with its own illustrated frame.
03Narrated and captioned. A soft storytelling voice, captions kept small so they do not fight the picture.
Voice:
Elli
Captions:
Subtle
Music:
Cinematic
Format:
Vertical 9:16
What comes out
The settings this page ships with, and the shape they produce
Every example on this page is labelled with the exact configuration behind it, including the one that says it is a drawing.
Layout mock-up
The lighthouse keeper who kept a light burning for a ship…
Style: Cinematic
Voice: Elli
Captions: Subtle
Ratio: 9:16
A mock-up of the frame and caption layout at this page’s settings — TikTok, Reels and Shorts, cinematic music bed. It is a drawing, not a rendered video: we have no published render library to show you yet, and a stock clip dressed up as our output would be worth less than saying so.
Your input
“The lighthouse keeper who kept a light burning for a ship that never came”
Plus this page’s brief, verbatim
Tell this as a short story in five beats: situation, complication, escalation, turn, resolution. Use past tense and concrete sensory detail. No moral at the end and no explaining the story after it has finished — the last line is part of the story.
That string is not marketing copy — it is the instruction the script writer is given before it reads a word of yours, and it is the only thing separating this page from the other forty-seven.
The shape that comes back
Scene 1 — the hook
One line that works with no context, because the viewer arrived mid-scroll.
Scenes 2–6 — the body
One idea each, one or two spoken sentences, each with its own image prompt and its own duration.
Final scene — the close
A line that lands the point rather than asking for a follow.
Per scene — the shot
The exact image prompt the model would receive, with this page’s style suffix already appended.
This is the structure, not a saved example. The box at the top of the page writes the real one — with your wording, in about ten seconds, free once you sign in.
5–8 scenes
each with its own line, its own frame and its own length
Under 60s
cut to the narration, not to a fixed per-scene timer
Word-timed
captions burned from the voice model’s own alignment
1080-class
vertical, square or landscape — no watermark on any plan
01
Five beats, because a short has room for exactly five
Under a minute, a story gets five moves: the situation, the thing that goes wrong, the escalation, the turn, and the landing. Any more and the beats get too thin to feel; any fewer and the turn arrives before the viewer cares about it.
That is the structure this preset asks for by name. The turn is the whole reason the video exists, so the pacing is arranged to arrive there with about ten seconds left rather than in the last two.
02
Illustration, not footage
Stories are the format where generated visuals are genuinely better than stock, because the scene you need has never been filmed. A lighthouse in the specific weather your story describes does not exist in a stock library, and a nearby lighthouse in different weather quietly undercuts the writing.
Each beat’s frame is generated from that beat’s own line, in a consistent style, which is closer to illustrating a book than to editing a video. The captions default to the subtle style here, small and out of the way, because on a story the picture is doing work the captions should not fight.
03
What it will not do
The same character will not look the same in every frame. Each image is generated independently, so a recurring person’s face drifts between beats.
Stories that survive this are about a situation, a place or an event, and describe people by role rather than by face — the keeper, the mapmaker, the crew. Stories that need a consistent protagonist in shot after shot are better served by something built for character consistency, which we are not pretending to be.
Who this is for
Four people this was built for
Storytelling channels
Anecdotes, folk tales and historical oddities, illustrated rather than filmed.
Authors and worldbuilders
A scene from the book as a short, made without commissioning an illustrator.
Brands with an origin story
The founding anecdote, told properly once, instead of buried on an about page.
Teachers
The story behind the fact, which is the part students actually remember.
Compared
What this replaces
Six things somebody has to arrange to make one short, and who arranges them.
Input
The usual way: A script, a storyboard, and an illustrator with a two-week queue.
With Veedio: A premise. The beats and the frames come back together.
Visuals
The usual way: Stock footage of somewhere that is nearly the place in the story.
With Veedio: A frame generated from the beat itself, so it is the place in the story.
Voiceover
The usual way: A narrator booked by the hour, re-booked for every rewrite.
With Veedio: A soft storytelling voice, re-recorded free whenever a beat changes.
Captions
The usual way: Auto-captions that drift, then an hour nudging keyframes.
With Veedio: Burned in from the voice model’s own word alignment, so they land on the syllable.
Turnaround
The usual way: Days, and most of it waiting on someone else.
With Veedio: About a minute, and you watch it fill in scene by scene.
Changing it
The usual way: Re-record the line, re-cut the timeline, re-export, re-upload.
With Veedio: Edit the line and regenerate that one scene. Regenerating a scene is free.
The usual way
With Veedio
Input
A script, a storyboard, and an illustrator with a two-week queue.
A premise. The beats and the frames come back together.
Visuals
Stock footage of somewhere that is nearly the place in the story.
A frame generated from the beat itself, so it is the place in the story.
Voiceover
A narrator booked by the hour, re-booked for every rewrite.
A soft storytelling voice, re-recorded free whenever a beat changes.
Captions
Auto-captions that drift, then an hour nudging keyframes.
Burned in from the voice model’s own word alignment, so they land on the syllable.
Turnaround
Days, and most of it waiting on someone else.
About a minute, and you watch it fill in scene by scene.
Changing it
Re-record the line, re-cut the timeline, re-export, re-upload.
Edit the line and regenerate that one scene. Regenerating a scene is free.
Related tools
Same engine, different preset
Twenty more pages over the same pipeline. Each one is a different brief, a different look and a different page — and each one prints its brief the way this one does.
A premise is enough — a situation and, if you have it, the turn. It will write the five beats. The more of the ending you supply, the more the story will be yours rather than the model’s.
Will the character look the same in every scene?
No. Frames are generated independently, so faces drift between beats. Stories that describe people by role rather than by face hold up; stories that need one consistent protagonist on screen do not.
Can I use it for Reddit-style story videos?
You can, with one caveat: there is no gameplay background here, so the visuals are illustrated scenes rather than footage under a voice. If the gameplay layer is the format, this is not it.
How long is a five-beat story?
Around 40 to 55 seconds. Each beat is one or two sentences, and each scene runs exactly as long as its narration.
Why are the captions small by default?
Because on a story the image is doing narrative work, and full-height bold captions cover it. The subtle style keeps the words available without competing. Both louder styles are one dropdown away.
Can I write the story myself and just have it produced?
Yes — use the script-to-video preset instead. It is briefed to keep your wording and only split it into scenes.
Does it handle dialogue?
Only as narrated dialogue: one voice reads the whole thing. There is no per-character casting, so a story that turns on two voices will not sound the way you are imagining it.
What about content restrictions?
The image models refuse graphic and explicit prompts, so a story that depends on those will come back with weaker frames. Horror and tension are fine, and there is a horror preset tuned for exactly that.
Do I have to sign up to try this?
Yes, a free one. Signup takes a moment and needs no card. Once you are signed in, the box at the top writes the script and the shot list for no credits — that is the part that decides whether the video is worth making — and rendering it into a file uses the 60 free credits the account starts with, about four videos.
What does it cost after the free credits?
Starter is $19 a month, Pro $49, Studio $149. A credit is a unit of pipeline cost rather than a video, so a five-scene short costs less than a nine-scene one, and you see the estimate before you spend anything.
Do I own the videos, and is there a watermark?
You own them, on every plan including the free credits, and there is no watermark on any plan, no resolution cap and no platform withheld. Four things do follow the plan, because they are what a video costs to make: AI motion and sound effects need Starter, a presenter avatar and the choice of video model need Pro, and a free account tops out at a minute where paid plans go to three. Everything else — every voice, every language, every caption style, every aspect ratio — is the same on all of them.