7 Tips That Turn ChatGPT 6 Astra Into AGI
One creator's seven-tip playbook for turning ChatGPT 6 Astra's computer-use and reasoning upgrades into real workflow gains.
September 7thAlex Finn runs both models through five self-designed benchmarks, crowns Opus 5 the winner on price and quality, then spends the back half explaining why he's not fully switching.
Claude Opus 5 outperforms Claude Fable 5 on cost, speed, and most custom benchmarks, but its verbose personality, tighter usage limits, and weaker coding harness keep the creator running multiple models instead of switching to Opus 5 entirely.
Alex Finn puts newly released Claude Opus 5 through five self-designed benchmarks against Claude Fable 5: a 3D roller coaster build, an Apple website clone, an agentic document scavenger hunt, a bug-fixing duel, and a bridge stress test. Opus 5 wins on visual quality, agentic reliability, and cost — roughly half the price of Fable 5 and usable at 100% of plan budget, where Fable 5 capped him at half. Despite the win, he flags three deal breakers: Opus 5's personality is verbose and had to be tamed by editing CLAUDE.md, Claude's usage limits still feel tighter than ChatGPT's, and its coding harness lags Codex and the ChatGPT app. His resulting stack: Opus 5 for hard problems, Fable 5 for planning, ChatGPT 5.6 as daily driver.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Create a free account →
Claims Opus 5 beats Fable 5, is cheaper and faster, but flags three deal-breaking weaknesses to come.

Lists the claims: beats Fable on almost every benchmark, half the price, full budget usable, significantly faster, questions whether there's any reason left to use Fable 5.

Introduces his own five-test benchmark suite built to compare the two models.

Both models build a 3D roller coaster simulator; Opus 5's is more detailed and cheaper (82 cents vs. roughly 50% more for Fable 5).

Both clone Apple.com from scratch; Opus 5's device recreations look closer to the real site than Fable 5's cut-off, distorted shapes.

Agentic document-search test across PDFs and spreadsheets; Opus 5 finds 5 of 8 items, Fable 5 stops partway over a false cybersecurity flag.

Debug Duel has both fix ~15 real open-source bugs (Opus 5 finishes ~25% cheaper); Breaking Point tests a bridge simulator under repeated car loads (Fable 5 holds slightly more weight but costs roughly double).

Final tally: Opus 5 wins on total cost, $6 versus $7.50 for Fable 5, declared the overall winner.

Personality sucks (verbose, unfocused, required a CLAUDE.md edit to tame), usage limits still feel tighter than ChatGPT's, and the Claude Code harness lags Codex and the ChatGPT app.

Recommends a three-model stack by task type and plugs an upcoming live bootcamp on Opus 5.
Benchmarks that beat another model on price and quality don't automatically win your daily workflow — personality, usage limits, and the surrounding tool can outweigh raw intelligence.
“Total cost was $6 for Opus, $7.50 for Fable five, and it's the winner.”
“The personality sucks. It absolutely sucks. I've never been so annoyed talking to a Claude model.”
“You ever have, like, that friend from high school who thinks he's just, like, way better than everyone else and way smarter than everyone else?”
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
Alex Finn opens by declaring Claude Opus 5 better than Fable 5 in almost every way — cheaper, faster, and the winner on his own five-test benchmark suite — before spending the video's back half walking through three things that stop him from fully switching.
A self-designed five-test suite the creator uses to compare coding models on visual output quality, agentic reliability, and cost.
The creator's resulting workflow after testing, splitting tasks across three different AI models by strength rather than picking one.
“I'm doing a boot camp on Opus five in an hour from me filming this. It's gonna be recorded. It'll be in the vibe coding academy. Link for that down below.”
Verbal plug placed right after the weakness deep-dive establishes credibility, paired with a description link and 'link down below' callout.
00:00
00:12
00:20
00:28
00:36
00:44
00:53
01:01
01:09
01:17
01:25
01:33
01:42
01:50
01:58
02:08
02:14
02:22
02:30
02:39
02:46
02:55
03:03
03:11
03:16
03:28
03:36
03:42
03:52
04:00
04:06
04:17
04:25
04:33
04:41
04:49
04:57
05:06
05:14
05:22
05:30
05:38
05:46
05:55
06:03
06:11
06:19
06:27
06:35
06:44
06:52
07:00
07:08
07:16
07:24
07:33
07:41
07:49
07:57
08:05
08:13
08:21
08:30
08:38
08:46
08:54
09:02
09:10
09:19
09:27
09:35
09:43
09:51
09:59
10:08
10:16
10:24
10:32
10:40
10:48Add Modern Creator as a preferred source and Google shows you more of our breakdowns in Search, Top Stories, and AI Overviews. It only changes what you see, and you can undo it in your Google settings anytime.
One creator's seven-tip playbook for turning ChatGPT 6 Astra's computer-use and reasoning upgrades into real workflow gains.
September 7thA benchmark-by-benchmark walkthrough of Anthropic's Fable 5.1 release, plus a custom test suite proving it beats GPT-5.6 Sol and demolishes Fable 5.
September 1stA tour of the open source, AI-first Linux desktop that lets you edit the operating system itself just by asking an agent.
August 28thA hands-on first look at the newly launched multi-agent AI product Grok Bot — its cloud-hosted agents, teachable skills, and agent-to-agent messaging — and whether it's good enough to replace open-source tools like Hermes and OpenClaw.
August 11thA creator's real-time demo of ChatGPT's new voice mode as an always-on dispatcher for AI agents across every device he owns.
July 27thA 12-minute breakdown of Sonnet 5 benchmarks, a live ChatGPT 5.5 head-to-head, a two-model workflow for Claude Code, and leaked Fable 5 strings.
June 30th