Claude Skill · Scale With Steven

Tenframe

Turn an idea into a 10-frame storyboard, then a finished AI video with a synced voiceover. Your character, setting, and real product stay consistent across every shot.

Script first → storyboard → video → audio. The workflow that actually ships.

The result — built end-to-end
🔊 Turn sound on — founder voiceover. 30s · 9:16 · claymation · one real product.
A second example: 2D cel explainer
🔊 30s · 9:16 · hand-drawn 2D cel · burned-in captions · the real product bottle.

Same pipeline, a completely different look. This one answers one curiosity question (what does a kid’s body actually do with the germs it meets), teaches the mechanism through a cleanup crew, and lands on the product. The bottle is passed in as a reference image so the label is the real one, and the captions are aligned word by word against the finished voiceover rather than typed by the video model.

Its storyboard
2D cel storyboard, ten frames
Ten frames, one per beat, built before a single second was rendered
How it works
01

Co-write the script

Tenframe gets the story first: angle, POV, the beats. It confirms aspect ratio + duration before a single pixel.

02

Build the storyboard

A 10-frame contact sheet, rendered with your real product photo as a reference so nothing gets invented. Sheets headed for video are built with Seedream, ByteDance’s own image model.

03

Render the video

The sheet drives Seedance 2.5, up to 30 seconds in a single pass. It’s the consistency anchor that stops the model from drifting shot to shot.

04

Finish the audio

A voiceover, fit to each beat and mixed over ducked ambient. One clean, on-brand spot.

The storyboard behind it
Storyboard Part 1 — the pain
Part 1 — the pain (10 frames, 15s)
Storyboard Part 2 — the solution
Part 2 — the solution (10 frames, 15s)
Why it works

The storyboard is the anchor

Ten loose shot descriptions drift. The character's face changes, the setting shifts. One contact sheet used as a reference holds continuity across the whole video.

Your real product, never invented

The skill passes your actual product (and founder) photo as a reference image. The picture guarantees fidelity. The words don't.

Story before pixels

It co-writes the script with you and gets sign-off first. The best ad here came from a real founder's origin story, not a generic template.

Reliable on any model

The fiddly parts (rendering, audio, assembly) are bundled scripts with the hard-won gotchas baked in, so it ships clean even on smaller models.

Faces stay locked

For a recurring person (a founder, a spokesperson), Tenframe builds a character sheet first: real face, every angle and expression, pupils leveled. That sheet rides along as the reference, so the face never changes mid-video.

The right image model for the job

Anything headed into Seedance is generated with Seedream, ByteDance’s own model. Keeping both halves in one family is what gets a reference past the likeness gate and transfers identity cleanly. Sheets from other generators get rejected.

Captions and sound, finished

Burned-in captions are aligned word by word against the final voiceover, with phrase breaks taken from your own script, because most of the feed watches muted. Music sits under the voice on a sidechain duck, so it drops when she speaks and lifts in the gaps.

The script is the product

A shared scriptcraft guide runs every format: one spine question, a hook that drops mid-reaction, teaching through friction instead of lecture, and a banned list of the phrases that make writing smell like AI.

Boards that show motion

Every panel encodes movement: motion arrows, camera-move labels, action lines, a subject that progresses across the frame. A row of centered portraits is a failed board, and Tenframe knows it.

Thirteen styles built in

Name the look. The whole pipeline adapts: sheet, prompts, render.

Premium 3D Claymation Realistic UGC POV Origami Paper-Craft Cinematic Slice-of-Life Exaggerated 3D Explainer POV Documentary 2D Cel Body-World NEW Chalkboard Explainer NEW Brick-Built Toy NEW In-Car Confessional NEW Two-Hander Podcast NEW

Five are brand new, and every style now ships with its named tell, the one giveaway that makes it read as AI, plus the cheap fix. Three carry a full pipeline of their own. Exaggerated 3D Explainer: a 9:16 vertical short that answers one curiosity question, oversized-head characters, 8 to 12 beats, built on an idea filter that kills weak topics before you spend a cent. POV Documentary: a first-person time-travel vlog where a host walks into a real historical era, talks to locals, and teaches, closing on the line people quote.

Get it

Drop it into your Claude skills folder and say “storyboard a 30-second video ad” or “make me a 3D explainer about why phones get hot.”

↓ Download the skill View source on GitHub →
# install (Claude Code)
unzip tenframe-skill.zip -d ~/.claude/skills/tenframe
# then set your keys (see SETUP.md)
export FAL_API_KEY="..."   # video
export OPENAI_API_KEY="..." # storyboard image + voiceover