Drop a voiceover; we transcribe it into timestamped lines.
StoryAnimator.app · illustrated story video studioPay as you go
Turn any voiceover into an illustrated story video
Make faceless YouTube videos from a voiceover — no camera, no face needed. Drop in an audio narration or record your own. StoryAnimator transcribes it into timestamped sentences, generates one consistent image per line, and lays it onto a real editing timeline synced to the waveform — ready to export as MP4.
8
Timestamped lines
0:58
Narration length
6
Image models
16:9
Export ratio
Contact
Questions, a bug, or a feature you wish existed — write to us and a human answers.
Example project — this is what StoryAnimator makes. It’s read-only: start your own project to edit and export.
1
Upload your audio
Drop a voiceover or narration file. We run speech-to-text and split it into sentences with start and end times.
Your narration
Drop audio or video here, or click to browse
Audio (WAV · MP3 · M4A · AAC · FLAC) or video (MP4 · MOV · WEBM) — up to ~35 min. A video's soundtrack is used automatically.
Auto-detect language · timestamped sentences
source.wav
No file yet
Waiting for audio…
Waiting
Upload an audio file to begin transcription.
Language
2
Review the transcript
Every sentence becomes a shot. Hover or click a row to highlight it — these timings drive image durations on the timeline.
Transcript & shots
transcript.txtread-only
shots.list0 shots
How many images?
drag or step — each shot becomes one generated image
0
fewermore
Story style
Pick one art style and set three guardrails — they apply to every image so the whole video feels like one film.
Art style
Pick a preset or describe your own — this single choice drives the whole look.
Applies to the cast and every generated frame.
Visual styleMood, medium, palette, lighting
Character bibleFixed designs reused every shot
Negative promptWhat to keep out of frame
Environment · optionalLeave empty — each shot follows the story. Fill only to pin one fixed setting everywhere.
Generation setup
Image modelWhich AI draws each frame
Video formatSets the shape of every image
Reference imageoptional
Upload character ref
Upload voice first to unlock this step.
Character cast
Generate a portrait of each character, refresh any you don't like, or edit a prompt with the pencil — the approved cast locks character consistency.
5
Results gallery
• Each generated 16:9 frame shows its index, filename, and the prompt that produced it.
• Export the whole job as a ZIP with manifest.
• Animating shots is optional — skip it and click Continue to export your video from the still images.
storyboardno job yet
Edit any shot's image prompt, then hit Generate. Re-roll individual frames anytime after. rendered
Tip: click a frame to select it · shift-click for a range · ⌘/Ctrl-click to add · right-click for Delete / Cut / Copy / Paste (paste asks before or after).
estimate
Video modelHow each still is animated into motion
How to animateDefault motion & feel for every shot — per-shot “🎬 Motion” overrides it
Animated: 0:00 · est. 0 credits
Image quality—
Nano Banana keeps characters consistent, accepts a reference image, and follows the per-sentence prompt.
$0.0000
0 frames @ $0.0000 · estimated total
$0.0000
0%
Transcribe audio before generating.
6
Video preview
Every frame laid end to end, each clip stretched to its sentence duration over the narration waveform. Scrub, preview, and export to MP4.
sequence.editno sequence
SHOT 01
PREVIEW
0:00.0 / 0:00.0
—
Shot timeline
0:000:00
Narration volume
Background musicNo music selected
Upload music you own, or licensed / Creative Commons music with attribution.
Subtitles
Subtitle styleKaraoke styles sync word-by-word to the voice (preview & full export)
Captions come from your timestamped transcript and stay in sync with each shot — no manual timing needed.
Import
Rebuild a video from existing assets — no regeneration. Three ways in:
① Video — split into frames + audiomp4, mov, webm
Splits the video into shot frames and lays its audio underneath. Leave Frames blank to auto-pick (~1 every 3s).
② Frames onlytimestamp filenames (e.g. 049.png)
Each frame is timed from its filename (the timestamp). Pair with a soundtrack below.
③ Soundtrack / audiomp3, wav, m4a…
Sits in the audio track underneath the frames. No re-transcription.
Reconstruct the project (reverse pipeline)subtitles + name + style
Transcribes the imported soundtrack, writes a caption onto each frame by its timing, and drafts the project name (if none), visual style, character bible, intro/outro & thumbnail — the reverse of the normal pipeline.
Export qualityformat set in Look
~ size shown here
Render
Ready · records in real time with audio
▶
YouTube details
Title, description, and tags — auto-written from your narration when the video is ready. Edit anything before you post.
▣ PUBLISH TO YOUTUBE
publish.meta
Title 0 / 100
Tags comma sep.
Description
These fill in automatically once your video finishes exporting — or click ✨ Generate with AI to write now.
💬
3-second hook
📌Put your most important link at the very top of the description.
👶In YouTube Studio choose “No, it's not made for kids” (unless it is) so reach isn't limited.
🎬Add an End screen in the last 15s pointing to another video — keeps viewers watching.
Open YouTube upload ↗Export your MP4 in the Editor, click Open YouTube upload, drop in the file, then paste your copied title / description / tags.
thumbnail.genfirst free · regen 1 credit
Click Generate to preview
Overlay text — ≤4 bold words
Image prompt — no text in image
1280 × 720 · ready to upload to YouTube
$
Credits and usage
Transcript and ZIP exports are free. Image generation and video export use credits.
account.ledger
total.bill
Transcription$0.0000
Character portraits—
Image generation$0.0000
Video render$0.0000
Total spent this session
$0.0000
Regular users see credits only. Admins see internal USD cost for margin tracking.
↺
History
Your saved projects. Re-open one to view its frames; re-rolling a shot replaces the stored image with the newest version.
history.log
Projects are saved automatically when you generate images — and updated whenever you re-roll a still or re-animate a shot, so this always shows your latest version.
StoryAnimator works in Firefox, but Chrome or Edge export faster — Firefox has to download a one-off 30 MB audio encoder on your first export.
voice.rec
Record and edit your voice
Record a take, select a range on the waveform, cut mistakes, undo or redo edits, then save it as the voice for this project.
📖 Reading script
Scroll to zoom · Shift+scroll to pan · Double-click to reset
1×
0:00.0
SPEED1.00×
BOOST×1
0:00.0 - 0:00.0
Session stays available until this page is refreshed.
voice.gen
Generate a voiceover
Paste your script and pick a voice — we'll narrate it and turn it into a story video. Charged in credits by length.
Pronunciation fixes (optional)Persian doesn't write short vowels, so a name or uncommon word can be guessed wrong. Say it here once and every voiceover in this project follows it.
0 / 5000
Pick a voice and paste a script to begin.
Add 1–3 clear samples (one speaker, no background music or noise) — upload files or record live. Long files are trimmed to the first 2 minutes; that is all cloning needs. Cloning creates a custom voice you can narrate with; the credit cost is shown on the button.
0:00
vocal.music
AI vocal music
Write lyrics (or let AI draft them), then turn them into a sung song — for music videos or children's nursery rhymes. The song becomes your project's voiceover.
Auto picks Suno for non-English lyrics, MiniMax for English.
music.gen
Generate background music
Describe the mood and style — we'll compose an original track to loop under your story. Charged in credits by length.
0:30
Slide 0:05–5:00 · loops for longer videos
Describe the music to begin.
Done
credits.shop
Top up credits
Buy StoryAnimator credits
Add any amount. Every credit is worth the same — bigger top-ups simply earn bonus credits.
Pay as you go · One-time · No subscription · Credits never expire
$
Pick an amount to continue to Stripe Checkout.
account.auth
Save projects · buy credits · export video
Sign in to StoryAnimator
Welcome back — sign in to keep your projects and credits.
best-results.guide
StoryAnimator · Workflow guide
How to get the best results
Voice — audio & script.Start from your own recording (WAV/MP3/M4A), record straight into the browser, or generate an AI voiceover. Whichever you pick, the audio is transcribed into timestamped lines — and every line becomes one shot. Fix any wording in the transcript, and use the “How many images?” slider to set how many shots the story splits into: more shots = shorter, snappier frames.
Look — style & cast.Pick the art style from the dropdown — that one choice owns the look of every frame. The visual-style and character notes are drafted for you and describe mood, palette and who your characters are; edit them freely. For any character you can generate a portrait or upload your own reference photo — the photo is read by AI, the description updates to match (skin tone, hair, wardrobe), and that identity carries through every shot.
Storyboard — prompts & generate.Each card shows a short editable “what to draw” line plus the caption that will appear on screen. Editing here is free — editing before you generate is what saves credits. Then hit Generate. Afterwards you can re-roll a single frame, upload your own image for one shot (upload icon), or bulk-upload a folder (UPLOAD button next to ZIP) to fill every shot in order. Right-click a frame for delete / cut / copy / paste. Animate is optional — shots that are already animated aren’t charged again.
Edit & publish.The timeline auto-syncs each frame to your waveform. Add intro/ending cards (auto-written for you), then export MP4. Your browser encodes it as fast as your machine allows, keeping animation, transitions, cards and subtitles. Browsers without that encoder fall back to recording in real time — there the ⚡ Fast export tick offers a quicker stills-only route (no animation, transitions or cards). Your YouTube title, description, tags and thumbnail are drafted from the narration, ready to copy.
Back up — download what you make.Treat the app as a workspace, not a vault: download your assets as you go and keep your own copy. Every shot has a download button (the clip if animated, otherwise the still), ZIP grabs all frames at once, the voiceover and the thumbnail each have their own download, and ⤓ Extract in History packs the whole project — media, audio and text — into a .zip you own. That .zip can be re-imported here any time, and anything you’ve downloaded can always be re-uploaded: your own audio, a reference photo, a single frame, or a whole folder of images.
History — one project at a time.Your project saves automatically as you generate, re-roll and re-animate, so History always holds your latest version — though only your most recent project is kept. Starting a new one takes its place, so run ⤓ Extract first whenever you’d like to keep the old one. As a matter of good practice, we recommend downloading anything you’re happy with as you go: a copy on your own machine is yours permanently, independent of your account or connection, and can be re-imported or re-uploaded here at any time. You can also re-import a project .zip, or switch History off from the History tab, which clears what’s stored there.
Tip: editing prompts before generating saves credits, and re-rolling or uploading lets you fix any single frame without redoing the whole set — and extract a .zip backup before you start anything new.
research.youtube
StoryAnimator · Find a video idea
Find a video idea
Start from a proven topic. Let AI suggest what's hot on YouTube, or search your own — you'll turn it into a script in a couple of steps.
or
or
Filters
These are the top videos on this topic. Pick a few you'd like yours to be like — the AI studies them and invents a fresh idea, not a copy.