The argument in one line.
Claude Fable 5 long-horizon focus makes it the first model capable of coordinating a complete multi-tool video production pipeline — script, voice, avatar, motion graphics, editing, and self-verification — from a single prompt without human oversight.
Read if. Skip if.
- You use Claude Code professionally and want to understand what the Mythos-class Fable 5 tier unlocks versus Sonnet or Opus.
- You are building AI-powered content production pipelines and want a real end-to-end example with real token costs.
- You work with ElevenLabs, HeyGen, or ffmpeg and want to see how Claude orchestrates them inside a /goal session.
- You are evaluating whether a dollar-200-per-month Max plan is worth it for agentic video production use cases.
- You want a step-by-step tutorial you can copy — the creator explicitly says results are not replicable without his pre-built Hyperframe skills.
- You are not already familiar with Claude Code's /goal command and skill system — this video assumes fluency with both.
The full version, fast.
Nate Herk used Claude Fable 5 and a single /goal prompt to generate a complete YouTube video — researched script, ElevenLabs voice clone, HeyGen avatar, ffmpeg editing, GSAP motion graphics — in about one hour. The model self-verified by rendering frames and re-rendering failures. Total cost: roughly 380K tokens, or about 40% of a dollar-200-per-month Max plan. Honest caveats: you cannot replicate this without pre-built Hyperframe skills, sub-agents used cheaper models for verification, and Sonnet is probably sufficient now that the pipeline exists.
Chat with this breakdown — free.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Create a free account →Where the time goes.

01 · The artifact plays
AI-generated video runs with on-screen overlays revealing AVATAR SYNTHETIC, VOICE CLONED, SCRIPT CLAUDE.

02 · Fable 5 capability overview
Mythos tier, Stripe 50M-line migration, Pokemon FireRed vision test, Slay the Spire long-horizon test, pricing.

03 · Pipeline walkthrough
Script via voice playbook, ElevenLabs chunks, HeyGen Avatar 5, Playwright workaround, ffmpeg stitch, GSAP graphics, self-verification.

04 · Fourth wall break
Creator steps out of the AI artifact and speaks directly — confirms the video was entirely AI-generated.

05 · Actual Claude Code session
Screen recording: ~380K tokens, 40% of /month plan consumed in one hour, full /goal prompt on screen, caveats about replicability.
Lines worth screenshotting.
- Claude Fable 5 used Playwright to click HeyGen's UI when the Avatar 5 API did not expose the model — browser automation is the escape hatch when APIs lag behind.
- Voice generation chunks must stay under 60 seconds or the cloned voice starts to drift — a constraint that shapes the entire pipeline architecture.
- The /goal prompt's most effective line was a reputation stake: 'you should only stop when you are 100% confident — it will damage my reputation.' Consequences outperform specifications.
- Self-verification via frame rendering — render, review visually, re-render failures — is the quality gate that makes fully unsupervised production viable.
- Sub-agents in the verification workflow ran on models below Fable 5; only the orchestrator was Mythos-class, which kept costs to 380K tokens.
- One hour of Fable 5 work consumed 40% of a dollar-200-per-month Max plan — the economics only work for content you would otherwise pay a human editor significant money to produce.
- The Mythos tier was locked to vetted security partners until Fable 5 — this video is a first-week stress test by a practitioner, not a controlled benchmark.
- Claude built every motion graphic card as live HTML animated with GSAP, then rendered it into the video — code as compositing tool, not just logic.
- Fable 5 beat Pokemon FireRed using only raw screenshots with no maps or navigation aids — older Claude models needed a full helper harness.
- A 50-million-line Ruby codebase migration that would have taken a team two months was compressed to a single day in the Stripe benchmark.
One prompt built a finished video — here is what the cost means.
The pipeline is real, but the 80-dollar session bill and pre-built skill dependency are the honest constraints behind the headline.
- Chunking TTS generation into sub-60-second segments prevents voice drift — a non-obvious constraint that shapes the entire pipeline architecture for any AI voiceover workflow.
- When an API does not expose a feature you need, Playwright browser automation is the escape hatch — Claude drove HeyGen's UI by hand until the API caught up.
- Self-verification via frame rendering — render frames, review visually, re-render failures — is the quality gate that makes fully unsupervised production viable without a human review loop.
- A reputation stake in the prompt ('it will damage my reputation') outperforms a specification list as a quality signal — giving the model context for why quality matters produces better judgment.
- Sub-agents handling verification ran on models below Fable 5; only the orchestrator was Mythos-class — a cost-management pattern worth copying for any long-horizon agentic workflow.
- One hour of Fable 5 orchestration consumed 40% of a 200-dollar-per-month Max plan — the economics only hold for content you would otherwise spend significant human time and money producing.
Terms worth knowing.
- Mythos class
- Anthropic's model tier above Opus, previously restricted to vetted security partners. Fable 5 is the first Mythos-class model on a standard paid plan.
- Hyperframes
- A Claude Code skill system for generating HTML/GSAP animated cards that render as video motion graphics inside ffmpeg compositions.
- /goal
- A Claude Code slash command that frames the session around a single outcome the model pursues autonomously, spawning sub-agents and tools until the goal is complete.
- Avatar 5
- HeyGen's newest motion engine for AI avatar video rendering, which was not initially accessible through HeyGen's public API at time of this video.
- Long-horizon focus
- The ability to maintain task coherence across millions of tokens — the primary improvement Anthropic cites for Fable 5 over Opus 4.
- Voice drift
- Degradation in voice clone quality when a TTS generation runs too long; mitigated by splitting scripts into sub-60-second chunks.
Things they pointed at.
Lines you could clip.
“I just typed one prompt into Claude Code and walked away. And everything else — the research, the script, the voice, the avatar, the motion graphics — all of it happened on its own.”
“This ate up about 40% of my a month plan. So in one hour, it ate up almost half of the plan.”
“I said, you should only stop when you are 100% confident that this is a high quality video. This will be going out to my YouTube channel. So if it doesn't look good, it will damage my reputation.”
Word for word.
Don't just watch it. Burn it in.
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
The bait, then the rug-pull.
What you're watching was not filmed. The avatar is synthetic, the voice is a clone, and every word of the script was written by Claude. The creator typed one prompt, walked away, and came back to a finished video he had never seen.
Named ideas worth stealing.
AI video production pipeline
- Script: Claude reads source, fact-checks, writes in creator voice
- Voice: ElevenLabs, chunked under 60s to prevent drift
- Avatar: HeyGen Avatar 5, Playwright workaround if API lacks it
- Edit: ffmpeg stitch plus word-level transcription
- Motion graphics: HTML/GSAP via Hyperframes, rendered into video
- Verification: frame render, visual review, re-render failures
Six-stage pipeline Claude orchestrated autonomously in one /goal session
How they asked for the click.
“if you guys enjoyed the video or learned something new, please give it a like”
Delivered by the AI avatar before the creator breaks the fourth wall — structurally clever because the AI delivered the CTA before the human revealed the trick











































































