Ink follows the voice
Every drawing is anchored to a phrase in the narration. The stroke starts a beat before its word so the ink lands as you say it — the detail that separates a whiteboard video from a slideshow.
Give it a prompt, a PDF, or your own voice. Dynamo.sh writes the script, composes each board, and animates a hand drawing it — synchronised to the word being spoken. Then it hands you the timeline to edit.
One video on us, then add your own Groq and Fish Audio keys — both free to create — and keep going for what the providers charge you, which is pennies. Or clone it and run the whole thing on your own machine, where there is no limit at all.
The commercial whiteboard-video tools are good. They are also subscriptions that meter your output, keep your source material, and hand back an mp4 you cannot open. This is the other trade.
Every drawing is anchored to a phrase in the narration. The stroke starts a beat before its word so the ink lands as you say it — the detail that separates a whiteboard video from a slideshow.
Upload a recording and the boards re-time to what you actually said. Add a transcript and the timing becomes exact. Or record on webcam and appear as an inset beside the visuals.
Drag elements on the canvas, drag bars on the timeline, or just ask in chat. Every change is a validated operation applied to a document you can read, diff and version.
Chalkboard, graph paper, kraft, ruled pad, sharpie. Each is a complete system — surface, stroke texture, palette and handwriting — not a colour swap on the same template.
Vertical for Shorts, horizontal for YouTube. Columns collapse into rows, flow diagrams run downwards, type resizes to the edge that runs out first. Not a letterboxed crop.
Groq, Anthropic, OpenAI or a local Ollama. Your own image API for photographs. Built-in system voices, or a cloud one. Swap any of them without touching the rest.
Name a thing and it is found — from a built-in set drawn stroke by stroke, or searched across open icon libraries that need no account. Only the word itself is sent, never your prompt or documents, and you can switch the lookup off entirely.
The board for scene one is settled the moment its narration is spoken, so there is no reason to wait for an encode of the whole thing. Scenes stream out as they are made — first one on screen in seconds — and what travels is the timeline rather than the pixels, so your own page draws it at its own resolution. Encode a file later, or never.
Working stacks down the board the way a teacher writes it — each line under the last, earlier lines staying put across scenes so the reader can look back up. The rule being applied sits in the margin, a tick or a cross goes on the line it judges, and the pen draws one thing at a time rather than three at once.
A sentence, a PDF, a pasted article — or all three.
It writes the narration and places every element on the canvas: position, size, colour, draw order.
Frame by frame, deterministically, straight into ffmpeg. No frame is ever guessed.
Chat, drag, or type exact numbers. Re-render costs one render, not one generation.
There is no per-video price here, because there is no per-video cost to us beyond a machine. What a video actually costs is the model that writes it and the voice that reads it, and both of those are billed to whoever owns the key. So that is exactly where the line is drawn.
Free
On our keys, on our machine. Enough to see whether the thing does what you need before you set anything up.
Create an accountYour own keys
Add a Groq key and a Fish Audio key on your account page — both free to create, a few minutes each. You pay the providers directly, which for a short explainer is pennies, and we do not charge credits for work done on your own keys.
How to get the keysNo limit
Clone it, run ./setup.sh, and it is yours: no accounts, no
allowance, no telemetry, and your documents never leave the machine.
An account takes a moment, and the first video is on us.
Or run it yourself and skip all of this — it is the same code either way.