The argument in one line.
The unlock with Fable 5 is not switching to it but routing each stage of your build to the cheapest model that still clears the bar, because effort tier and model choice are two separate dials most users never touch.
Read if. Skip if.
- You use Claude Code and are on a Fable 5 subscription, or are planning to upgrade.
- You have burned through weekly credits faster than expected and want a systematic fix.
- You run multi-agent or dynamic workflows and need a model-routing strategy across planning, execution, and verification stages.
- You want a concrete cheat sheet for when to use Fable max vs medium vs Sonnet vs Opus 4.8.
- You want benchmark comparisons against GPT, Gemini, or other providers -- the video explicitly skips benchmarks.
- You do not use Claude Code or the Anthropic Claude platform.
The full version, fast.
Fable 5 is the most capable model Anthropic has shipped, but treating it as a daily driver will exhaust your credits in days. The central insight: Anthropic already auto-downgrades Fable to Opus 4.8 for sensitive requests, routing by risk. You apply the same mechanic, routing by cost. Use Fable at high or max only where decisions compound -- planning and final verification. Run all execution volume through Opus 4.8, Sonnet, or local models. Switch mid-session via /model so the spec stays in the thread. At medium effort Fable already beats Opus 4.8 at max, so you rarely need the top tier for anything but the plan.
Chat with this breakdown — free.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Create a free account →Where the time goes.

01 · With great power comes token burn
Hook framing: Fable 5 is real, but so is the cost. The premise is set without benchmarks.

02 · The real trap -- June 22 meter and addiction
Anthropic announcement on screen: Fable 5 exits flat-rate subscription June 23 and becomes metered. The trap is forming a dependency before the meter hits.

03 · System prompt x-ray: Fable 5 vs Opus 4.8
Leaked system prompt compared against Opus 4.8. 80 percent identical. Five new rules around self-harm and life sciences. Filler-word instructions removed -- suggesting retraining.

04 · Anthropic routes down for safety, you route down for cost
Fable auto-downgrades to Opus 4.8 on cybersecurity and health requests. Steal that mechanic: route bulk work to cheaper models voluntarily.

05 · When to use each tier of effort (the Goldilocks zone)
Core framework. Fable medium beats Opus 4.8 max. Effort tier is a separate dial from model choice. Planning and shipping = high/max. Execution = medium or lower, Sonnet, or Opus.

06 · The tactical loop: plan expensive, /model, execute cheap
Concrete mechanic: run Fable on max for planning, create a spec file, type /model to switch to Sonnet or Opus for execution, then re-invoke Fable to probe edge cases.

07 · Three real recipes: marketing site, 3D website, CRM
Three worked examples. Marketing site: Fable high plan, Opus med build, Fable low verify. 3D website: Fable x-high plan, Opus plus Sonnet agents build, Fable high verify. CRM: Fable max plan, x-high dynamic workflows, high-model verify.

08 · Don't be tribal: mixing models and providers
Brief section against model loyalty. Fable plus Codex via OpenAI plugin extension. Use whatever clears the task at lowest cost.

09 · The one workflow to walk away with
Consolidated: Fable max/high planning, orchestrate execution with agents using skills/MCPs/CLIs, Fable verification, ship, iterate. The paradigm persists as model names change.

10 · Fable vs Opus, tier by tier
Comparison: Fable low ties Opus 4.8 max. Fable medium beats it. Fable max crushes it.

11 · How to use Fable 5 responsibly
Wrap. Benchmarks are manufactured; results are what matters. Fable short-circuits on cybersecurity. Opus is still more reliable day-to-day. CTA for free cheat sheet and community.
Lines worth screenshotting.
- Fable 5 at medium effort already beats Opus 4.8 at maximum effort -- throttling down costs less and still wins.
- Anthropic auto-downgrades Fable to Opus for cybersecurity and health requests; you can apply the same routing logic to every bulk task you run.
- 80 percent of the Fable 5 system prompt is identical to Opus 4.8 -- new rules are almost entirely around self-harm and life sciences safety categories.
- Filler-word instructions were removed from the Fable 5 system prompt, suggesting those habits were retrained into the base model rather than injected at runtime.
- Most Claude Code users never touch the effort dial -- they pick a model and leave effort at default, leaving easy token savings on the table.
- Using /model mid-conversation lets you shift from Fable to Sonnet without losing context -- the spec lives in the thread, not in the model.
- Fable still short-circuits on cybersecurity-adjacent requests even when benign and above board, making Opus the safer daily driver for now.
- Sub-agents in a dynamic workflow should almost never run on Fable; the orchestrator sets the plan, cheap models do the volume.
- Verification stage scales with risk: marketing site gets Fable low, 3D website gets Fable high, CRM gets high model for auth and integration checks.
- Planning is where decisions compound -- it is the only stage that earns max-effort spend. Execution volume does not.
- Benchmarks can be manufactured to grab headlines; the only signal that matters is whether the output meets your specific task bar.
- You can combine Fable with Codex via the OpenAI plugin extension -- model loyalty is a budget liability, not a virtue.
Model choice and effort tier are two different decisions.
Picking the smartest model is only half the decision -- the effort dial is equally powerful and almost universally ignored.
- Power and cost scale together -- Fable 5 is genuinely better, but every token costs more, so defaulting to the best model for everything is unsustainable.
- Anthropic is moving Fable 5 to metered billing after a flat-rate trial window -- forming a dependency before the meter starts is the trap to avoid.
- 80 percent of the Fable 5 system prompt is identical to Opus 4.8, with additions concentrated in self-harm and life sciences safety rules.
- Filler-word instructions were removed from the prompt, suggesting those tendencies were retrained into the model rather than patched at the prompt layer.
- Anthropic already auto-downgrades Fable to Opus for risky asks -- routing by task type, not by habit, is the transferable lesson.
- Fable 5 at medium effort already beats Opus 4.8 at maximum effort -- throttling down costs less and still wins.
- Effort tier is a separate dial from model choice; planning and shipping earn high/max, execution volume earns medium or lower.
- The /model command lets you downgrade mid-session without losing context -- use it after the spec is locked to shift execution to a cheaper model.
- After execution, re-invoke Fable to probe edge cases rather than trusting the cheaper model to self-certify.
- Complexity of the plan determines how high you go on the first Fable call; sub-agents always use cheaper models regardless of project complexity.
- Verification scales with integration risk -- the more moving parts that can silently fail, the higher the model tier needed to spot them.
- Provider loyalty is a cost liability -- Fable and Codex can run side by side; use whichever clears the specific task at the lowest tier.
- The paradigm of plan-expensive / execute-cheap / verify-smart will persist as model names change -- it is the routing logic that matters, not the specific models.
- Fable low ties Opus max, Fable medium beats it, Fable max crushes it -- even a heavily throttled Fable run outperforms a fully maxed Opus run.
- Benchmarks are manufactured for headlines; only your specific task results matter.
- Fable still short-circuits on cybersecurity-adjacent requests even when benign, making Opus the safer default for daily coding.
Terms worth knowing.
- Effort tier
- A setting within Claude Code (low / medium / high / max) that controls how much compute the model spends reasoning before answering. Distinct from the model choice itself.
- Tokenomics
- The economics of token consumption -- how cost scales with input/output length, model tier, and frequency of use.
- Dynamic workflow
- A Claude Code pattern where an orchestrator model spins up a variable number of sub-agents at runtime based on the task.
- Ultracode
- The highest-tier agentic execution mode in Claude Code, combining dynamic workflows with tool-use orchestration.
- Metered access
- Usage-based billing where each token consumed charges separately, as opposed to a flat-rate subscription with a usage cap.
- /model command
- A slash command in Claude Code that lets you change the active model mid-conversation without starting a new session, preserving all prior context.
Things they pointed at.
Lines you could clip.
“The average person will run out of credits by the time they say good morning to Fable.”
“Benchmarks don't matter. They can be doctored. They can be manufactured. They grab the headlines, but the only thing that matters are your results.”
“Fable five on low is still a very competent model.”
Word for word.
Don't just watch it. Burn it in.
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
The bait, then the rug-pull.
With great power comes an even greater token bill. Unlike most AI tool videos that lead with benchmarks, this one leads with the credit limit. The title is a warning, the hook is a price tag, and everything that follows is a routing strategy.
Named ideas worth stealing.
The Effort Tier Matrix
- Plan: Fable high/max
- Orchestrate: Fable medium or Opus
- Execute: Sonnet / Opus / local
- Verify: Fable low to high (scales with risk)
- Sub-agents: always cheaper models
Effort and model are two separate dials. Most users pick a model and never touch the effort dial.
The Routing Mechanic (steal from Anthropic)
Anthropic auto-downgrades risky asks from Fable to Opus 4.8. Apply the same mechanic for cost: route high-stakes planning to Fable, route volume execution down to cheaper models.
Three Recipe Cards
- Marketing site: Fable high plan / Opus med build / Fable low verify
- 3D website: Fable x-high plan / Opus plus Sonnet agents build / Fable high verify
- CRM: Fable max plan / x-high dynamic workflows / high-model verify
Complexity of the project drives how high you go on planning; sub-agents always use cheaper models; verification scales with integration risk.
How they asked for the click.
“Check out the first thing down below -- early adopters community and free cheat sheet in the second link.”
Double CTA pattern -- community link and free lead magnet. Community pitch appears twice at 7:45 and 13:57.





































































