Script free with an account·render in about a minute·no watermark on any plan
AI Listicle Video Generator
A topic in, a five-beat numbered list out, with one generated frame per item and the count on screen.
Start free — no credit card, and the script and the shot list cost no credits.
01Name the list. The subject and, if you care, the number. “Five” is the default because five items and a hook fit under a minute.
02One item, one scene. Each item gets its own line, its own frame and its own length, so the count is visible in the edit rather than only in the words.
03Read and captioned. A bright read with word-by-word highlight captions, which is what makes a number land as it is spoken.
Voice:
Bella
Captions:
Highlight
Music:
Uplifting
Format:
Vertical 9:16
What comes out
The settings this page ships with, and the shape they produce
Every example on this page is labelled with the exact configuration behind it, including the one that says it is a drawing.
Layout mock-up
Five things nobody tells you about renting your first flat
Style: Documentary
Voice: Bella
Captions: Highlight
Ratio: 9:16
A mock-up of the frame and caption layout at this page’s settings — TikTok, Reels and Shorts, uplifting music bed. It is a drawing, not a rendered video: we have no published render library to show you yet, and a stock clip dressed up as our output would be worth less than saying so.
Your input
“Five things nobody tells you about renting your first flat”
Plus this page’s brief, verbatim
Write this as a numbered listicle. Open with a hook that names the list and the payoff, then give each item its own scene: the number said out loud, the item stated plainly, and one sentence that justifies it. Keep every item concrete — a thing, a number, an action — and finish on the strongest one rather than the weakest.
That string is not marketing copy — it is the instruction the script writer is given before it reads a word of yours, and it is the only thing separating this page from the other forty-seven.
The shape that comes back
Scene 1 — the hook
One line that works with no context, because the viewer arrived mid-scroll.
Scenes 2–6 — the body
One idea each, one or two spoken sentences, each with its own image prompt and its own duration.
Final scene — the close
A line that lands the point rather than asking for a follow.
Per scene — the shot
The exact image prompt the model would receive, with this page’s style suffix already appended.
This is the structure, not a saved example. The box at the top of the page writes the real one — with your wording, in about ten seconds, free once you sign in.
5–8 scenes
each with its own line, its own frame and its own length
Under 60s
cut to the narration, not to a fixed per-scene timer
Word-timed
captions burned from the voice model’s own alignment
1080-class
vertical, square or landscape — no watermark on any plan
01
Why a listicle is the easiest short-form shape to get right
A list video has a structure the viewer can hold in their head from the first line. They know how many things are coming, they know roughly when it ends, and every scene change is a promise being kept. That is why the format survives every algorithm change: it answers the two questions a scrolling viewer asks — what is this, and how long.
It is also the shape that survives being cut. If item four is weak you delete it and the video still works, which is not true of a story or an argument. That makes a listicle the safest format to publish on a schedule, and it is why so many channels that post daily post lists.
The pipeline is built for that: each scene is a separate line with its own generated frame and its own duration, so removing one item in the editor removes exactly one beat and nothing else has to be re-timed.
02
Five, not ten, and the reason is structural
This page writes to the default minute, which is five to eight scenes. A ten-item list would spend every one of them on an item and leave nothing for a hook or a close, which is the difference between a video and a slideshow.
So this preset writes five. One hook, five items, one closing line, and the whole thing reads in about forty-five seconds. If you want ten, the honest answer is that it should be two videos — and the second one usually performs better because the hook can promise the half you did not publish yet.
03
What makes an item worth a scene
Generic items are the failure mode. “Be consistent” is not an item; “post at the same time every day for a month and then look at your retention graph” is, because it contains an action and a check.
The brief pushes the writer towards things you can picture, which matters twice over here: a concrete item gives the image model something to draw, so the frame behind item three is a specific object rather than another abstract gradient.
Who this is for
Four people this was built for
Daily-posting channels
A shape you can fill with a different subject every day without the video feeling like the same video.
Newsletter and blog writers
You already write lists. This is the version of one that gets watched instead of skimmed.
Product and tools accounts
Five picks, five frames, one voice — the format the whole recommendation genre runs on.
Anyone testing a niche
A list is the cheapest way to find out which of five angles your audience actually wanted.
Compared
What this replaces
Six things somebody has to arrange to make one short, and who arranges them.
Input
The usual way: Write the list, then storyboard it, then find five clips that match five items.
With Veedio: The subject. The numbering, the pacing and the frames come out of it.
Visuals
The usual way: Five stock clips that half-match, or five slides in a template.
With Veedio: A frame generated per item, from that item’s own line.
Voiceover
The usual way: Book a VO artist, or record yourself and cut the breaths out.
With Veedio: Five ElevenLabs voices, re-recorded free whenever the line changes.
Captions
The usual way: Auto-captions that drift, then an hour nudging keyframes.
With Veedio: Burned in from the voice model’s own word alignment, so they land on the syllable.
Turnaround
The usual way: Days, and most of it waiting on someone else.
With Veedio: About a minute, and you watch it fill in scene by scene.
Changing it
The usual way: Cutting item four means re-timing the countdown and re-recording the numbers.
With Veedio: Delete the scene. The rest keeps its own timing, and regenerating a scene is free.
The usual way
With Veedio
Input
Write the list, then storyboard it, then find five clips that match five items.
The subject. The numbering, the pacing and the frames come out of it.
Visuals
Five stock clips that half-match, or five slides in a template.
A frame generated per item, from that item’s own line.
Voiceover
Book a VO artist, or record yourself and cut the breaths out.
Five ElevenLabs voices, re-recorded free whenever the line changes.
Captions
Auto-captions that drift, then an hour nudging keyframes.
Burned in from the voice model’s own word alignment, so they land on the syllable.
Turnaround
Days, and most of it waiting on someone else.
About a minute, and you watch it fill in scene by scene.
Changing it
Cutting item four means re-timing the countdown and re-recording the numbers.
Delete the scene. The rest keeps its own timing, and regenerating a scene is free.
Related tools
Same engine, different preset
Twenty more pages over the same pipeline. Each one is a different brief, a different look and a different page — and each one prints its brief the way this one does.
Not in one video at this length. A minute is five to eight scenes, so a ten-item list would have no room for a hook or a close. Five items is the shape that fits, and splitting a ten into two videos usually performs better anyway — though a 3-minute video in the composer does have room for ten if you want it.
Does it count down or up?
Up by default — item one first — because the strongest item last is what holds a short-form viewer. If you want a countdown, say “count down from five” in the box and the script comes back reversed.
Will the number appear on screen?
It is spoken and it appears in the captions as it is said, because the captions are burned from the voice’s own word timings. There is no separate animated number overlay — that is renderer work we have not built.
Can I supply the items myself?
Yes. Paste your five items into the box and the script writer will keep them, write the hook and the connective tissue around them, and turn each into a scene.
What subjects make bad lists?
Anything where the items are not really separable. A process is a tutorial, not a list, and forcing it into numbers makes the middle scenes feel arbitrary — the tutorial preset is the right page for that.
Do I need an account to see the script?
Yes, a free one, and signup needs no card. The box at the top writes the list, the scene split and the shot list for no credits. Rendering it into a video is what spends them, starting from the 60 the account comes with.
Do I have to sign up to try this?
Yes, a free one. Signup takes a moment and needs no card. Once you are signed in, the box at the top writes the script and the shot list for no credits — that is the part that decides whether the video is worth making — and rendering it into a file uses the 60 free credits the account starts with, about four videos.
What does it cost after the free credits?
Starter is $19 a month, Pro $49, Studio $149. A credit is a unit of pipeline cost rather than a video, so a five-scene short costs less than a nine-scene one, and you see the estimate before you spend anything.
Do I own the videos, and is there a watermark?
You own them, on every plan including the free credits, and there is no watermark on any plan, no resolution cap and no platform withheld. Four things do follow the plan, because they are what a video costs to make: AI motion and sound effects need Starter, a presenter avatar and the choice of video model need Pro, and a free account tops out at a minute where paid plans go to three. Everything else — every voice, every language, every caption style, every aspect ratio — is the same on all of them.