Academy video series
A model-agnostic AI video production system, customised for the LayerV Academy. It runs with Claude, ChatGPT, open-source models and any other generative model. Built from scratch around the Academy in under 3 weeks: your lessons, your character, your brand, your cost structure. It runs without its creator, is fully customised, and is meant to evolve with the team.
Contents
The Academy lessons were already written. What did not exist was a repeatable route from a written lesson to a finished film.
The economics of a video series are decided by a few structural choices. Here is what those choices saved, and how each figure was established.
| Choice | What it saves | How it was established |
|---|---|---|
| The system itself | $35 with it against $348 without | A ratio close to 10:1 that holds at any provider's price point |
| Reusable shot library | $30.48 already saved, against $25.50 of footage generated | Reuse has returned more than the footage cost. Measured, and revised downward from $48.73 by an audit |
| Generate at 720p, composite type locally at full resolution | 56 to 60% off every generated shot | Measured on tight crops rather than assumed: no visible loss on the character, and none at all on the type |
| Revisions rebuild locally | No generation cost, however many rounds it takes | Type, panels, captions, transitions and sound are never generated, so changing them cannot cost credits |
The compounding part. Reuse has already returned more than the entire footage spend, and that happened while generation was deliberately held to the strict minimum. The saving grows on its own once the character palette is settled, because every shot then stops being a one-off and becomes stock.
4 stages, each with an explicit contract: what it reads, what it writes, what it may not touch, and where it stops. The chain halts wherever a human decision belongs, and nowhere else.
| Stage | What it does | Waits for | Spends |
|---|---|---|---|
| 1. Script | Lesson page to manifest and script, with burned captions and their sync timings | Script approval | No |
| 2. Storyboard | Manifest to shot list and prompts, rendered as real images with teaching panels composed at the true geometry of the film | Storyboard approval | No |
| 3. Produce | Approved storyboard to stills, footage, voice and finished file | Written approval of an exact amount | Yes |
| 4. Revise | Notes on a finished cut to a local rebuild, with before and after at each note | Nothing | No |
It can be run by an operator who has never seen it. An
operator with no prior knowledge of the chain fed a lesson in at one
end and took a finished, voiced, captioned episode out at the other.
The command layerv-academy-video makes the workflow run
end to end.
The first cut is a draft, not a delivery. No generative pipeline lands the intended result on the first pass, and this one is no exception: expect rounds on pacing, emphasis and wording before an episode is right. That is precisely what the architecture is for. Those rounds rebuild locally, they cost no generation, and they are the cheapest part of making the film.
3 of the 4 stages cost nothing to run. A lesson can be written, storyboarded, reviewed, rejected and rewritten as often as needed. Money enters once, at a single stage, after an exact figure has been approved.
Each stage is a skill with a written contract. What makes them reliable is what they do and what they refuse to do: ownership of each field is assigned to exactly 1 skill, so no stage can silently overwrite another's work.
Every guardrail below exists because something went wrong at least once. Each was proved able to fail before it was trusted: run against the broken code first, watched to fail, then adopted.
This is version 1 of a system shaped around the Academy as it exists today. It was built to be changed by the team that runs it, not frozen at handover.
Every word the viewer reads arrives by one of 2 routes, and the choice belongs to the operator. It is worth seeing the difference, because it decides both what a change costs and what the letters look like.
| Layer | What it does | Changing it |
|---|---|---|
| The assistant | Reads the method files and writes the script, the storyboard and the revision notes | The rules are plain markdown, not code, so they are readable by any capable assistant. Operated so far with a single one |
| The assembly machinery | Turns approved material into a finished film: cuts, transitions, type, teaching panels, captions, sound mix, loudness | Calls no AI at all. Standard code and standard media tooling. This is why revisions cost no generation, and why no vendor sits between you and a finished film |
| The generation models | Produce the footage, the stills and the voice | Named in configuration. The registry declares 8 video endpoints and 3 voice options; change the identifier and the spend gate reprices automatically from the live rate |
Today the series is generated with the frontier Seedance video model, named in configuration rather than written into the code.
What exists today: The method files and the framework reference page are already in your hands.
The system is in service, and like any version 1 it has a roadmap. What follows is the open list as the system itself keeps it, rather than a list reconstructed for this report.
| Open item | Why it matters | Owner |
|---|---|---|
| Character palette not approved | Every shot generated before it is settled stays provisional. A palette change would invalidate the $25.50 library. This is the single decision that unlocks the reuse economics at scale | Design lead |
| First run on the consolidated spend gate | The 4 approval steps were merged into 1 command. That consolidated path has not yet carried a paid run through to a finished film, so the first one is worth walking step by step on the cheapest episode | Operator, with the reviewer approving the amount |
| Prompt grammar migration | 3 of the 4 existing episodes still carry the earlier grammar. A guard refuses those shots by name, so the gap surfaces before money moves rather than after | Operator |
| Two prompt-level checks still to build | The validator confirms that a figure is spoken somewhere in the episode, but not which card it sits on. It also does not yet refuse a prompt that reserves space while scripting a light effect across it. Both gaps have produced errors a check would have caught | Operator |
| Loudness of the earliest films | The first films play 4 to 5 dB quieter than the current chain produces. Levelling them is free; redelivering is an editorial call | Content reviewer |
Once the team has run an episode or 2, 2 pieces are worth building on top of what exists.
The rule above the others. Nothing is generated without a written authorisation naming the exact amount, calculated at the moment of asking. A revision is not a generation: revising a finished cut is local, costs no credits, and takes minutes.
That distinction is the entire economics of the series. Holding it is what keeps the cost per episode where it is.
No. The approval names an exact amount and works as a ceiling, checked to the cent at submission. Over it, the run refuses and says why.
Rarely, and that is true of any generative pipeline. Plan on a few rounds of notes on pacing, emphasis and wording. The system was built around that expectation: those rounds rebuild locally and cost no generation, so iterating is the cheap part.
Write your notes the way you would say them. The rebuild works through composited layers and settings rather than a new generation, and takes a few minutes rather than a new invoice. A note that genuinely needs new footage is named as such and priced separately, never absorbed silently.
No generation: scripts, storyboards, every caption, panel, transition and sound mix, and all revisions. Paid: new footage, new stills, new voice lines. Nothing else.
No. The chain runs on documented commands, stops at every point where a decision belongs to you, and takes review notes in plain language.
Through the prompt. Each shot's motion comes from a written prompt, and how carefully that prompt is written is what decides how Vega moves and presents. The system supplies the structure, the brand constraints and the automatic refusals; the craft of the prompt itself belongs to the operator, and it is where the difference between an average shot and a good one is made.
Yes, in the method file of the stage that applies them. A change there applies to every future episode, and the automated checks protect the rest of the series while you change it.
Machine time varies with how many shots are generated versus reused, and with the provider's queue: roughly 15 to 45 minutes from an approved storyboard to a finished file is a realistic range. The calendar time on top of that is review: how fast the script and the storyboard get approved, which is where the schedule actually lives.
No, and that is deliberate. The assembly machinery calls no AI at all, so nothing proprietary sits between a lesson and a finished film. The rules each stage follows are plain markdown rather than code, so any capable assistant can read them; the series was authored with Claude, and nothing in the chain depends on that choice. The video model and the voice are named in configuration, so either can be swapped for a more capable or a cheaper one without touching the chain.