How Claude Fable 5.1's Text Watermark Actually Works (And How to Get Around It)
It's not a hidden mark in the text, it's a loaded die on word choice, and only a key Anthropic won't hand out can read it.
September 2ndA screen-by-screen read of Anthropic's Opus 5 announcement, benchmark chart by benchmark chart.
Anthropic's newly announced Claude Opus 5 matches or beats the rival Fable 5 model across nearly every benchmark category while costing roughly half as much per task, and the jump from Opus 4.8 to Opus 5 is a bigger leap than the gap between competing labs' flagship models.
Anthropic's Opus 5 announcement shows the model matching rival Fable 5 at high and extra-high effort settings while costing about half as much per task, and only trailing at max effort by roughly 0.5% on CursorBench. The bigger story per the host is the jump from Opus 4.8 to Opus 5 itself — a larger leap than the gap to competing labs — driven by stronger self-verification on long-horizon agentic tasks: the model checks its work against a goal and iterates rather than one-shotting. Anecdotal examples include writing its own computer-vision pipeline to reconstruct a machine part from a drawing it couldn't view, fixing a bug at its root cause rather than the symptom, and building its own test harness to validate a market data feed with no live reference to check against. Pricing lands at $5/$25 per million input/output tokens, unchanged from Opus 4.8 and about half of Fable 5's rate.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Create a free account →
Comparison table across agentic terminal coding, knowledge work, novel problem-solving, agentic search, multidisciplinary reasoning, computer use, agentic coding, business workflows, legal, health, biology. Opus 5 beats Fable 5 and Opus 4.8 in most categories, trails only on multidisciplinary reasoning. CursorBench: within 0.5% of Fable 5 at max effort, at half the cost; roughly equal at high/extra-high effort.

The Opus 4.8-to-Opus 5 jump is framed as the bigger story than the Fable 5 comparison. Unlike Sonnet 5 (described as disproportionately expensive per task despite improvements), Opus 5 is token-efficient relative to both Sonnet 5 and Fable 5.

Two embedded Anthropic demo clips: Opus 5 generating a wind-tunnel airflow visualization over a car model, and building an interactive 3D cell artifact. Framed as closing a longstanding gap since Claude models can't natively generate images.

Three anecdotal examples from Anthropic's page: reconstructing a machine part in FreeCAD from a drawing it couldn't view (wrote its own CV pipeline), fixing an open-source bug at its root cause instead of the symptom, and building its own test harness to validate a market data feed with no live reference. Misaligned-behavior audit and vulnerability-finding also improved.

Cybersecurity classifiers are less restrictive than Fable 5's (85% less likely to trigger) via a Cyber Verification Program for vetted users. Opus 5 is priced at $5/$25 per million input/output tokens — same as Opus 4.8, about half of Fable 5.
Benchmark wins depend heavily on effort level and cost framing, and the real story in this announcement is a self-verification jump, not just a leaderboard position.
“Claude Opus five just dropped in. Anthropic is claiming we now have a model that is getting close to Fable five outputs at half the cost.”
“At max max effort, Fable five does pull ahead, but most people aren't actually operating at max. They're either operating at high or extra high.”
“It went ahead and wrote its own computer vision pipeline to pull the geometry from raw pixels and then reconstructed the whole machine part.”
“If the numbers are true, we kind of almost have a better Fable for half the cost.”
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
Anthropic just published its Opus 5 announcement, and the host reads it live on screen — a benchmark table and a run of cost-per-task charts that put the new model within half a percent of rival Fable 5's peak score, at roughly half the price per task.
00:00
00:04
00:08
00:11
00:14
00:17
00:20
00:24
00:27
00:30
00:33
00:37
00:40
00:43
00:46
00:49
00:53
00:56
00:59
01:02
01:06
01:09
01:12
01:15
01:19
01:22
01:25
01:28
01:31
01:35
01:38
01:41
01:44
01:48
01:51
01:54
01:57
02:00
02:04
02:07
02:10
02:13
02:15
02:20
02:23
02:26
02:30
02:31
02:36
02:39
02:42
02:46
02:49
02:52
02:55
02:58
03:02
03:05
03:08
03:11
03:15
03:18
03:21
03:24
03:28
03:31
03:34
03:37
03:40
03:44
03:47
03:50
03:53
03:57
04:00
04:03
04:06
04:09
04:13
04:16Add Modern Creator as a preferred source and Google shows you more of our breakdowns in Search, Top Stories, and AI Overviews. It only changes what you see, and you can undo it in your Google settings anytime.
It's not a hidden mark in the text, it's a loaded die on word choice, and only a key Anthropic won't hand out can read it.
September 2ndA hands-on walkthrough of Impeccable, the open-source Claude Code design skill, and its new Live Mode and Worlds features for turning AI-generated web design from generic to genuinely good.
August 4thKimi K3's benchmark charts and rock-bottom per-token price look like a knockout blow to Claude and GPT — until a blind three-way build test and a real cost-per-task tally tell a much closer story.
July 17thA breakdown of three automation buckets — sales, research, and content — built on Claude Code and an indexed Obsidian vault, reclaiming five to ten hours a week.
July 15thA creator walks through five concrete levers — effort level, model delegation, token-saving skills, research offloading, and advisor mode — for keeping Claude Code costs and weekly usage caps under control.
July 3rdEarly access to OpenAI's next flagship model turns into a benchmark massacre, a string of jaw-dropping 3D demos, and one very ugly story about a model that lied about finishing a PR.
September 4th