Script free with an account·render in about a minute·no watermark on any plan
Explainer Video Generator
One idea, explained properly in forty seconds — the claim, the mechanism and why it matters.
Start free — no credit card, and the script and the shot list cost no credits.
01Name the concept. One thing. If it needs an “and”, it is two videos and both will be better for it.
02It builds the explanation. Claim, mechanism, example, consequence — the structure explanations actually follow.
03Check it before it renders. The script comes back first, which is where you catch the one sentence that is subtly wrong.
Voice:
Rachel
Captions:
Bold
Music:
Lo-fi
Format:
Vertical 9:16
What comes out
The settings this page ships with, and the shape they produce
Every example on this page is labelled with the exact configuration behind it, including the one that says it is a drawing.
Layout mock-up
How a heat pump can produce more energy than it consumes
Style: 3D render
Voice: Rachel
Captions: Bold
Ratio: 9:16
A mock-up of the frame and caption layout at this page’s settings — TikTok, Reels and Shorts, lo-fi music bed. It is a drawing, not a rendered video: we have no published render library to show you yet, and a stock clip dressed up as our output would be worth less than saying so.
Your input
“How a heat pump can produce more energy than it consumes”
Plus this page’s brief, verbatim
Write a short explainer. Structure it as: the claim that sounds wrong, the mechanism that makes it true, one concrete example with a real number, and why it changes what the viewer does. No jargon unless it is defined in the same sentence. One concept only.
That string is not marketing copy — it is the instruction the script writer is given before it reads a word of yours, and it is the only thing separating this page from the other forty-seven.
The shape that comes back
Scene 1 — the hook
One line that works with no context, because the viewer arrived mid-scroll.
Scenes 2–6 — the body
One idea each, one or two spoken sentences, each with its own image prompt and its own duration.
Final scene — the close
A line that lands the point rather than asking for a follow.
Per scene — the shot
The exact image prompt the model would receive, with this page’s style suffix already appended.
This is the structure, not a saved example. The box at the top of the page writes the real one — with your wording, in about ten seconds, free once you sign in.
5–8 scenes
each with its own line, its own frame and its own length
Under 60s
cut to the narration, not to a fixed per-scene timer
Word-timed
captions burned from the voice model’s own alignment
1080-class
vertical, square or landscape — no watermark on any plan
01
Explanations have a shape, and it is not a list of facts
A good explainer moves through four beats: a claim that sounds wrong, the mechanism that makes it true, a concrete example with a real number in it, and the consequence for the person watching. Skip the mechanism and it is trivia. Skip the consequence and it is a lecture.
That structure is what this preset enforces. It is also why the videos come out shorter than people expect — most of the length in a human-made explainer is the throat-clearing that this structure has no room for.
02
Abstract subjects and the picture problem
The hardest part of an explainer is that its subjects are usually invisible. There is no photograph of an interest rate, a protocol handshake or a supply chain.
The image prompt for each scene is therefore written as a literal, concrete scene rather than as the abstraction itself — the thing on the desk, the building, the object being held — which is what a human editor does when they cut to a b-roll shot. The 3D render style is the default here because clean, lit, uncluttered frames carry a diagram-like idea better than a photographic one.
03
The script comes back before anything is generated
Explainers are the format where being subtly wrong costs the most. A slightly mangled mechanism is worse than no video, and it is the thing a language model is most likely to produce.
So the script arrives first, in full, with every spoken line visible, and writing it costs none of the credits your free account starts with. Read it as if a domain expert were watching over your shoulder, fix the line that overstates, and only then spend anything on rendering it.
Who this is for
Four people this was built for
SaaS marketers
The feature nobody understands, explained in the feed rather than in a doc nobody opens.
Educators
The bit of the lesson students always ask about, as a standalone forty seconds.
Internal comms
A process change that a memo cannot carry and a meeting cannot fit.
Anyone answering the same question weekly
Make it once, link it forever, stop typing the same paragraph.
Compared
What this replaces
Six things somebody has to arrange to make one short, and who arranges them.
Input
The usual way: A script, a storyboard, an animator and three rounds of revisions.
With Veedio: The concept in one line, with the structure applied for you.
Visuals
The usual way: Motion graphics, quoted per second, changed only when it is worth re-quoting.
With Veedio: A generated frame per scene, in a clean style, regenerated free when you rewrite a line.
Voiceover
The usual way: Book a VO artist, or record yourself and cut the breaths out.
With Veedio: Five ElevenLabs voices, re-recorded free whenever the line changes.
Captions
The usual way: Auto-captions that drift, then an hour nudging keyframes.
With Veedio: Burned in from the voice model’s own word alignment, so they land on the syllable.
Turnaround
The usual way: Days, and most of it waiting on someone else.
With Veedio: About a minute, and you watch it fill in scene by scene.
Changing it
The usual way: A factual correction after delivery means a new animation pass.
With Veedio: Fix the sentence, regenerate that scene, keep the rest of the video.
The usual way
With Veedio
Input
A script, a storyboard, an animator and three rounds of revisions.
The concept in one line, with the structure applied for you.
Visuals
Motion graphics, quoted per second, changed only when it is worth re-quoting.
A generated frame per scene, in a clean style, regenerated free when you rewrite a line.
Voiceover
Book a VO artist, or record yourself and cut the breaths out.
Five ElevenLabs voices, re-recorded free whenever the line changes.
Captions
Auto-captions that drift, then an hour nudging keyframes.
Burned in from the voice model’s own word alignment, so they land on the syllable.
Turnaround
Days, and most of it waiting on someone else.
About a minute, and you watch it fill in scene by scene.
Changing it
A factual correction after delivery means a new animation pass.
Fix the sentence, regenerate that scene, keep the rest of the video.
Related tools
Same engine, different preset
Twenty more pages over the same pipeline. Each one is a different brief, a different look and a different page — and each one prints its brief the way this one does.
It is a language model, so treat the first draft as a competent draft and not as an authority. That is exactly why the full script is returned before any asset is generated — read it, fix what is wrong, then render.
Can it explain my own product?
Yes, if you supply the facts. Describe what it does and what it replaces in the box, or paste a link to your documentation. It cannot know anything about a product it has never seen.
Are there diagrams or animated charts?
No. Visuals are generated photographic or rendered frames, not data visualisations. If a chart is the point of the explanation, this is the wrong format.
How long should an explainer be?
Under a minute here, which is five to eight scenes. That is enough for one mechanism explained properly, and not enough for two — the constraint usually improves the video.
Can I control the reading level?
Say so in the box: “explain it for someone who has never studied economics” changes the vocabulary and the examples the scriptwriter reaches for.
What visual style suits explainers?
The 3D render style is the default because clean, lit frames carry an abstract idea well. Documentary works better when the subject is physical, and cinematic when it is historical.
Can I use my brand colours?
Not as a setting. You can push a look through the image prompts — every scene’s prompt is editable in the editor — but there is no brand kit, and pretending otherwise would waste your afternoon.
Can I make a series of them?
Yes, and it is the best use of the format. Keep the voice and style fixed, change the concept, and the set reads as one series rather than five one-offs.
Do I have to sign up to try this?
Yes, a free one. Signup takes a moment and needs no card. Once you are signed in, the box at the top writes the script and the shot list for no credits — that is the part that decides whether the video is worth making — and rendering it into a file uses the 60 free credits the account starts with, about four videos.
What does it cost after the free credits?
Starter is $19 a month, Pro $49, Studio $149. A credit is a unit of pipeline cost rather than a video, so a five-scene short costs less than a nine-scene one, and you see the estimate before you spend anything.
Do I own the videos, and is there a watermark?
You own them, on every plan including the free credits, and there is no watermark on any plan, no resolution cap and no platform withheld. Four things do follow the plan, because they are what a video costs to make: AI motion and sound effects need Starter, a presenter avatar and the choice of video model need Pro, and a free account tops out at a minute where paid plans go to three. Everything else — every voice, every language, every caption style, every aspect ratio — is the same on all of them.