The argument in one line.
Claude Opus 5 matches or beats Fable 5 on real one-shot coding tasks while costing half as much per token, splitting a two-test head-to-head 1-1 on raw output quality but winning decisively on price.
Read if. Skip if.
- You use an AI coding agent (Claude Code, Cursor, or similar) and are deciding which model to default to for build tasks.
- You care about cost per token as much as raw output quality when picking a coding model.
- You're curious what a single, unedited one-shot prompt produces across competing frontier models.
- You're already on Claude Max or Pro and weighing when Opus 5 beats sticking with Opus 4.8.
- You want a rigorous, multi-prompt statistical benchmark rather than one creator's two anecdotal tests.
- You don't use AI coding tools at all.
The full version, fast.
A creator sends the identical prompt to Fable 5, Opus 4.8, and the newly released Claude Opus 5: first, build a data-driven website about the post-AI economy; second, build an endless flight simulator over a city. Opus 5's economy site is dramatically more polished than both competitors, with a working 3D globe and accurate animated stats. On the flight simulator, Fable 5's city is denser and more detailed, with sound, so it wins that round. The tests end 1-1 on quality, but Opus 5 is priced at $5 per million input tokens and $25 per million output tokens — the same as its predecessor Opus 4.8, and half of what Fable 5 costs — so the creator's conclusion is to default to Opus 5 for daily agentic coding work.
Chat with this breakdown — free.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Create a free account →Where the time goes.

01 · Opus 5 launches at half Fable 5's price
Cold open recaps Anthropic's Opus 5 release: $5 per million input tokens and $25 per million output tokens, the same price as Opus 4.8 but half of Fable 5. The host calls out that Opus 5 is the new state of the art per Anthropic's own claims, with one exception: it stays behind on cybersecurity tasks.

02 · The case for switching to Opus 5 daily
The host argues there's little reason to keep defaulting to Fable 5 when Opus 5 offers comparable intelligence for half the cost, framing this as the emerging consensus in the agentic coding space, and asks viewers what further model comparisons they'd want to see.

03 · Test 1: the same prompt, three models, Fable 5 goes first
The rules are set: the identical prompt — build a web app about the state of the economy after AI — goes to Fable 5, Opus 4.8, and Opus 5. Fable 5's one-shot build (finished fastest, in about 7 minutes) has a polished 3D background and stat sections, but the host flags a visibly misaligned 3D element as unexpected from a frontier model.

04 · Opus 4.8's version: solid but not close
Opus 4.8's build of the same prompt has similar animated sections and 3D touches, but the host says it's 'nowhere close' to how good Fable 5's build looks.

05 · Opus 5's build: 'a different league'
Opus 5's one-shot economy site draws the strongest reaction of the video: accurate statistics, a professional-looking layout and fonts, an interactive 3D globe visualizing where frontier AI research and compute cluster geographically (Bay Area, Beijing, Bengaluru, London), and animated bar charts with glow effects the host calls better than both other models.

06 · Test 2: the flight simulator — Fable 5 wins this one
Second head-to-head: an endless flight simulator flying over a city. Fable 5's build has a denser, more colorful skyline with lights, tall structures, and working sound. Opus 5's version, while functional with plenty of buildings, is described as visually more basic and less detailed — evening the overall score to 1-1.

07 · Verdict: split on quality, decided on price
The host recaps both tests, lands on the two models being roughly equal in output quality across different task types, then repeats Opus 5's pricing and Anthropic's benchmark comparison as the deciding factor — half the price of Fable 5 for comparable performance — before closing with a request for viewer input on future model comparisons.
Lines worth screenshotting.
- Claude Opus 5 launched priced identically to Opus 4.8 — $5 per million input tokens, $25 per million output tokens — which is half the per-token cost of Fable 5.
- Anthropic's own release notes place Opus 5 at the new state of the art on its benchmarks, with one stated exception: it remains behind on cybersecurity tasks.
- Given the same single prompt, three models produced visibly different one-shot builds: Fable 5's had a 3D element that was visibly misaligned, Opus 4.8's was serviceable but plain, and Opus 5's was described as 'a different league.'
- Opus 5's economy-site build included a working 3D globe showing frontier AI research clustering in specific regions — Bay Area, Beijing, Bengaluru, London — pulled from a single prompt with no follow-up iteration.
- On a second, harder test — an endless-flight city simulator — the result flipped: Fable 5 produced the more detailed, more colorful city with working sound, and Opus 5's version was comparatively basic.
- Across two different one-shot build categories, the two models split 1-1 on output quality, suggesting neither is a strict upgrade over the other on raw capability alone.
- The deciding factor the creator lands on isn't which model 'wins' a given test — it's that Opus 5 delivers comparable results for half of Fable 5's price.
- Fable 5's one-shot economy-site build was the fastest of the three to complete, at roughly seven minutes.
How to actually compare two AI coding models
A single one-shot prompt run through competing models can flip winners between task types, so the real decision usually comes down to price for comparable output, not a single test's winner.
- The same prompt can produce very different results across models and across task types — a model that wins on one build type (data-driven website) can lose on another (physics-heavy game simulator).
- Don't judge a coding model from a single one-shot test; run it against more than one kind of project before deciding it's better or worse than an alternative.
- When two models produce roughly comparable output quality, price per token becomes the practical tiebreaker for which one to default to.
- Read a vendor's own benchmark claims for the caveats, not just the headline: Anthropic's own release notes flagged Opus 5 as behind on cybersecurity tasks specifically, even while claiming state-of-the-art elsewhere.
- A model that finishes a build fastest isn't necessarily the one producing the highest-quality result — speed and polish are separate axes worth checking independently.
Terms worth knowing.
- One-shot build
- Giving an AI coding model a single prompt and evaluating whatever it produces without follow-up corrections or iteration.
- Agentic coding
- Using an AI model inside a tool that can write, run, and modify code autonomously across multiple steps, rather than just answering a single question.
- Vibe coding
- Building software by describing the desired outcome in natural language and letting an AI model generate the implementation, with little to no hand-written code.
Things they pointed at.
Lines you could clip.
“Anthropic just released Opus five... and the results will shock you.”
“This is absolutely a different league.”
“Fable five and Opus five are quite equal in a lot of different ways... but the key difference here is the pricing.”
Word for word.
Don't just watch it. Burn it in.
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
The bait, then the rug-pull.
Anthropic just released Opus five, and in this video, we've vibe coded applications with Opus five and then built the same applications with Fable five and Opus four eight.
How they asked for the click.
“If you guys wanna see more tests, maybe comparing Opus five to different models... just let me know in the comment section down below and I'll release that in just a few hours.”
Soft, low-pressure ask for comments to steer a follow-up video rather than a product or subscribe pitch.









































































