How Claude Fable 5.1's Text Watermark Actually Works (And How to Get Around It)
It's not a hidden mark in the text, it's a loaded die on word choice, and only a key Anthropic won't hand out can read it.
September 2ndClaude Sonnet 5.5 launched today claiming to beat Opus 5.5 on a coding benchmark at half the token cost, so this AI YouTuber pulls up Anthropic's own announcement page and checks the numbers live.
Claude Sonnet 5.5 closes Anthropic's mid-tier gap by beating Opus 5.5 on select coding benchmarks at roughly half the token cost, but the gains flatten at maximum reasoning effort, so reading the cost-adjusted chart matters more than the headline score.
Anthropic released Claude Sonnet 5.5 promising 30% faster, 30% cheaper, and stronger across the board than Sonnet 5, which had been criticized as an expensive token hog. The host walks Anthropic's own charts: on Terminal-Bench 4.0, Sonnet 5.5 at max effort actually edges out Opus 5.5 at a lower cost per attempt, and token pricing (cache writes, input, output) is half of Opus's. But the gains aren't uniform — on FrontierCode, pushing to max effort scores worse than high effort while still costing more, so the practical takeaway is to default to high or extra-high effort rather than always maxing out. Sonnet 5.5 also jumps on knowledge-work and computer-use benchmarks and picks up the same cybersecurity and biology safeguards as Opus, which silently downgrade flagged high-risk prompts to older, weaker fallback models.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Create a free account →
The host states Anthropic's own framing: Sonnet 5.5 is 30% faster, 30% cheaper, and stronger across the board than Sonnet 5, following last week's well-received Opus 5.5 release.

Anthropic's benchmark table shows Sonnet 5.5 far outscoring Sonnet 5 on agentic coding (Terminal-Bench 4.0, FrontierCode, CursorBench), with the host noting Sonnet 5 was a known token hog on hard tasks.

The host argues raw scores don't matter without cost context, then walks the Terminal-Bench 4.0 accuracy-vs-cost chart (Sonnet 5.5 beats Opus 5.5 at max effort for less), the FrontierCode cliff where max effort scores worse than high effort while still costing more, and the per-million token pricing table showing Sonnet 5.5 at half of Opus 5.5's cache-write, input, and output rates.

On the GDPVal-AA real-world knowledge-work benchmark, Sonnet 5.5 lands near Opus 5.5 and about 400 points above Sonnet 5, with improvements also called out in computer use and chart/visual recognition.

Anthropic claims Sonnet 5.5 is a much better conversational partner than Sonnet 5 was, echoing the fix seen in Opus 5.5. The host then covers the safeguards section: strong cybersecurity capabilities trigger an automatic fallback to an older, weaker model on flagged high-risk prompts, the same applies to biology, and Anthropic is adding anti-distillation protections.

The host closes by framing Sonnet 5.5 as Anthropic's chance to finally compete in the cheap everyday-task tier where it had been losing ground, then points viewers to his own Claude Code course in the pinned comment.
A benchmark chart only tells you something useful once you read it against its cost per attempt and the effort tier it was run at.
“It's going to be 30% faster, it's gonna cost 30% less, and it's gonna be stronger across the board.”
“I wouldn't be surprised when you actually start using this thing for real in your day-to-day tasks that you find that max doesn't always mean you're getting a better outcome.”
“Sonnet 5 was actually a complete token hog, especially if you put it on very difficult tasks.”
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
Anthropic dropped Claude Sonnet 5.5 today, and this AI YouTuber pulls up the announcement page live to check the headline claim: 30% faster, 30% cheaper, and — on one benchmark — beating Opus 5.5 outright.
Judge a benchmark result by accuracy divided by cost per attempt, not by the raw accuracy number alone, since a higher score at a much higher cost isn't necessarily the better outcome.
“make sure to check that down in the pinned comment”
A single low-key mention at the very end, pointing to a pinned comment link rather than an in-video overlay or hard sell.
00:00
00:06
00:10
00:14
00:18
00:22
00:27
00:31
00:35
00:39
00:43
00:48
00:52
00:56
01:00
01:04
01:08
01:13
01:17
01:21
01:25
01:29
01:33
01:38
01:42
01:46
01:50
01:54
01:58
02:03
02:07
02:11
02:15
02:19
02:24
02:28
02:32
02:36
02:40
02:44
02:49
02:53
02:57
03:01
03:05
03:09
03:14
03:18
03:22
03:26
03:30
03:35
03:39
03:43
03:47
03:51
03:55
04:00
04:04
04:08
04:12
04:16
04:20
04:25
04:29
04:31
04:37
04:41
04:45
04:50
04:54
04:58
05:02
05:06
05:11
05:15
05:19
05:23
05:27
05:31Add Modern Creator as a preferred source and Google shows you more of our breakdowns in Search, Top Stories, and AI Overviews. It only changes what you see, and you can undo it in your Google settings anytime.
It's not a hidden mark in the text, it's a loaded die on word choice, and only a key Anthropic won't hand out can read it.
September 2ndA 10-minute screen-share walkthrough of the Anthropic announcement: what Fable 5 and Mythos 5 actually are, what they cost, and what the classifier guardrails really block.
June 9thA screen-by-screen read of Anthropic's Opus 5 announcement, benchmark chart by benchmark chart.
July 24thA 15-minute breakdown of the two-part feature that lets Claude spawn hundreds of isolated sub-agents for complex tasks and why the default single-session approach fails.
June 8thAn 8-minute first look at Anthropic’s new visual design tool — what it does, how it compares to Stitch and Lovable, and why the visual layer matters even when the code underneath is identical.
April 17thHiggsfield split its API from its up-to-$220/month plans, priced it below Fal, and shipped a free skill that lets Claude Code or Codex call it directly.
September 17th