One creator ran two frontier AI agents through the same 15 real work tasks and tracked the winner, the time, and the exact dollar cost for every single one.
Across 15 real work tasks run through both agents, GPT-6 Astra won 10 and cost $186 less overall, but Fable 5.1 still wins on wordy, detail-dense writing and one-shot design fidelity, proving the two tools have different strengths rather than one clear champion.
Who This Is For
Read if. Skip if.
READ IF YOU ARE…
You're paying for more than one frontier AI subscription and trying to decide which agent to reach for on which task.
You run an agency, course, or content business and want to see real cost-per-task numbers instead of marketing claims.
You're curious how far agentic AI has come on messy, real-world deliverables like tax P&Ls, subscription audits, and browser-driven course uploads.
SKIP IF…
You're looking for a tutorial on how to prompt either model — this is a results comparison, not a how-to.
You only care about coding benchmarks — none of the 15 tasks here are pure software engineering.
TL;DR
The full version, fast.
The video runs GPT-6 Astra and Fable 5.1 through 15 identical real-world tasks — a branded slide deck, sales copy, a tax P&L, an email subscription audit, meeting analysis, an event recap video, a sizzle reel, two 3D games, an eval SaaS, a knowledge-graph brain visual, an HTML explainer, browser-driven Canva and Skool tasks, a website clone, and a YouTube channel review — grading each on output quality plus the exact time and dollar cost. Astra won 10 of 15, driven by stronger structured/tabular work, browser automation, and vision tasks, while Fable won on wordier, detail-heavy writing and one first-pass design clone. Total spend was $513.36 for Fable versus $326.98 for Astra — Astra was $186 cheaper but took about 1 hour 44 minutes longer across all 15 runs combined. The conclusion: neither model dominates outright, so match the model to the task rather than picking one exclusively.
Free for members
Chat with this breakdown — free.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Cold open introduces the format — 15 real tasks (research, sales copy, taxes, subscriptions, meetings, video recaps, games, an eval SaaS, browser use, web design, YouTube analysis) will be scored on output, time, and cost.
01:14 – 03:35
02 · Use case 1: McKinsey-style SMB deck
Both models research and design a branded slide deck for a fake consulting brand. Fable's deck is wordier but more structured and branded (35 slides); Astra's is less wordy but feels less professional. Fable wins the output; Astra is less than half the cost ($12 vs $26) and 14 minutes faster.
03:35 – 06:05
03 · Use case 2: sales letter copy
Both write sales copy for a certification program using the creator's own research and second brain. Fable writes roughly 2,800 words with FAQs and a tuition section; Astra writes about 1,300 words and feels more high-level. Fable wins the copy but costs nearly 3x more.
06:05 – 08:15
04 · Use case 3: personal taxes
Both build a tax P&L and forecast across two quarters. Astra asks seven clarifying questions first and pulls a 3,739-row transaction ledger; its structure feels more trustworthy. Astra wins the output despite costing more ($22.64 vs $13.29) and taking longer (40 vs 22 minutes).
08:15 – 11:29
05 · Use case 4: subscription audit
Both scan years of email across two accounts to find every recurring subscription. Fable's spreadsheet leaves one tab completely empty and costs almost $50; Astra's version is complete and costs about $12.50. Astra wins clearly on both quality and cost.
11:29 – 13:44
06 · Use case 5: meeting analysis
Both mine leadership meetings for pain points and the highest-value automation to build. Astra analyzes 79 meetings versus Fable's 58, and Fable's run inexplicably costs $46.28 for six minutes of active agent time. Astra wins on both insight quality and cost ($5.62 vs $46.28).
13:44 – 17:36
07 · Use case 6: AIS Live recap video
Given 150GB of raw event footage and a vague prompt, both models cut an energetic 60-second recap video with music. Both outputs impress the creator; Astra edges out on live-clip usage and gets the slight win. Astra: 30 min / $16.10. Fable: 50 min / $26.05.
17:36 – 19:51
08 · Use case 7: sizzle reel
Both create a motion-graphic sizzle reel for the AIS brand from a reference video. Astra pulls real screenshots from the actual courses and adds 3D depth; the creator gives it the win on energy and creativity despite similar cost and time.
19:51 – 21:46
09 · Use case 8: 3D pistachio escape game
Both build a playable 3D platformer game with real physics. Both are completed in similar time (~43-46 seconds to win); Astra's version feels more premium and smoother to control, earning a narrow, subjective win.
21:46 – 24:49
10 · Use case 9: eval SaaS build
Both build a mini SaaS for scoring automations against a golden dataset and generating reports. Astra's interface feels more premium and less intimidating; it also costs half of Fable's spend despite taking longer in wall-clock time (2h25m vs 1h20m). Astra wins.
24:49 – 27:14
11 · Use case 10: Herk Brain knowledge graph
Both build a local visual knowledge graph of the creator's business notes and connections. The creator prefers Fable's simpler node layout over Astra's more overwhelming interface — the one use case where interface taste, not raw capability, decides the winner.
27:14 – 29:26
12 · Use case 11: HTML explainer doc
Both turn a YouTube tutorial into an annotated, screenshot-heavy HTML explainer doc for a beginner to learn from. Fable takes more, better-annotated screenshots and is judged the better learning experience, though Astra is markedly cheaper.
29:26 – 31:20
13 · Use case 12: browser-use painting in Canva
Both must open Canva, use the tools, and recreate a reference painting purely through browser control. Fable's result is described as unusable; Astra clearly wins both on output and cost.
31:20 – 32:53
14 · Use case 13: Skool course upload via browser use
Both must log into Skool and draft a new course module with video, description, and key points using only browser automation. Fable fails to upload the video entirely; Astra completes the task correctly. Astra wins decisively.
32:53 – 34:53
15 · Use case 14: one-shot website clone
Both are shown a coffee-brand site with heavy scroll animation and asked to build a similar-feeling site for a fictional brand with no other guidance. Fable's version matches the reference site's feel more closely and wins, despite the creator generally rating Astra higher on design.
34:53 – 37:04
16 · Use case 15: YouTube 12-month review
Both analyze a year of YouTube channel data and produce a strategy report. Astra's report is judged easier to read and more visual; Astra wins but costs about 3x more ($26.63 vs $8.53) and takes longer (42 vs 17 minutes).
37:04 – 38:35
17 · Final tally and verdict
Astra wins 10 of 15 use cases to Fable's 5. Total spend: Fable $513.36 over 9h35m, Astra $326.98 over 11h19m — Astra $186.38 cheaper but 1h44m slower overall. The creator leans toward Astra day-to-day but stresses GPT 5.6 Sol still covers most knowledge work, and both subscriptions stay in rotation.
Atomic Insights
Lines worth screenshotting.
Fable 5.1 tends to just start working on a prompt, while GPT-6 Astra asks clarifying questions first — on a tax task, that extra questioning alone increased confidence in Astra's output before any numbers were compared.
On a branded McKinsey-style SMB slide deck, Astra finished in 23 minutes for $12 versus Fable's 37 minutes for $26 — less than half the cost and 14 minutes faster — yet Fable's output still won on formatting quality.
For sales copy on a certification program, Fable wrote roughly 2,800 words to Astra's roughly 1,300, and the extra length (FAQs, tuition section, who-it's-for) was judged more persuasive despite costing nearly 3x more ($4 vs $1.43).
On a personal tax P&L across two quarters, Astra pulled a full transaction ledger of 3,739 rows and won on structure, but cost $22.64 versus Fable's $13.29 and took 40 minutes versus 22.
Auditing years of email for recurring subscriptions, Fable's output came in at nearly $50 and left one entire tab completely empty, while Astra did the same job for about $12.50 with no missing data.
Analyzing 58 leadership meetings, Fable's run somehow cost $46.28 for just six minutes of active agent time — a cost the creator could not fully explain even after checking session logs and usage stats.
Astra actually processed more source material on the meeting-analysis task (79 meetings analyzed versus Fable's 58) while still costing roughly a tenth as much ($5.62 vs $46.28).
Given a 150GB folder of raw event footage and only a vague prompt, both models independently produced a coherent 60-second energetic recap video with music, showing agentic video editing has crossed a real usability threshold.
For browser-driven tasks specifically — painting in Canva and uploading a course to Skool — Astra was rated clearly better, matching the creator's broader experience that GPT/Codex models currently outperform Claude-based agents at browser use.
Building a two-sided AI evaluation SaaS (test cases, golden datasets, run history, reports), Astra's version cost half of Fable's spend while taking longer in wall-clock time, and was preferred on design polish.
On a self-hosted 'brain' knowledge-graph visualization, the creator preferred Fable's simpler node layout over Astra's, showing win rate isn't just about raw model capability — interface taste matters just as much for a visual deliverable.
One-shot cloning a coffee-brand website's scroll animations, Fable's result matched the original site's feel more closely even though the creator generally rates Astra higher for pure design work.
Across all 15 tasks, Astra won 10 and Fable won 5, but the total cost gap ($186.38) came almost entirely from a handful of outlier-expensive Fable runs, not from Fable being systematically pricier on every task.
Total combined runtime across all 15 tasks was 9 hours 35 minutes for Fable versus 11 hours 19 minutes for Astra — Astra was cheaper in dollars but slower in wall-clock time overall.
The creator's stated takeaway is that GPT 5.6 Sol remains 'so, so good' for most day-to-day knowledge work, and the Astra-vs-Fable comparison is about edge cases and preference, not proof that either older/cheaper model is obsolete.
Takeaway
Match the model to the task, not the other way around
WHAT TO LEARN
Across 15 identical real-world tasks, the two agents split roughly along a line between structured/quantitative work and long-form written or subjective-design work, so the win column depends far more on task type than on either model being universally better.
02Use case 1: McKinsey-style SMB deck
On a branded slide deck, the wordier, more structured, more heavily-branded output won the presentation-deliverable test even though it cost twice as much to produce.
03Use case 2: sales letter copy
For persuasive sales copy, more words weren't padding: the version with FAQs, a tuition section, and a clear who's-it-for breakdown was judged more trustworthy to a paying prospect.
04Use case 3: personal taxes
A model that asks clarifying questions before starting and then shows its full supporting data (a multi-thousand-row transaction ledger, for a tax task) reads as more trustworthy than one that just hands over a recommendation.
05Use case 4: subscription audit
Cost outliers happen: one run cost nearly $50 and left a spreadsheet tab completely empty, while a comparable run on the same task cost a quarter as much with no missing data — always sanity-check a single run's cost before trusting it.
06Use case 5: meeting analysis
A model can burn ten times the dollar cost of its competitor for the same number of minutes of active work, so token usage and wall-clock time are not the same signal — check both.
07Use case 6: AIS Live recap video
Given only a vague creative prompt and a huge pile of raw source material, both tested agents could independently produce a coherent, usable short video with music — a bar that used to require a human editor.
08Use case 7: sizzle reel
In a subjective creative task, going out and grabbing real supporting assets (actual screenshots of the product being promoted) reads as more convincing than pure motion-graphic invention.
09Use case 8: 3D pistachio escape game
Even in a low-stakes creative build like a game, small differences in physics feel and control smoothness are enough to decide a close, subjective contest.
10Use case 9: eval SaaS build
A cheaper, faster-feeling agent can win the underlying build while still losing on interface polish, showing that build cost and finished-product taste are separate axes worth grading separately.
11Use case 10: Herk Brain knowledge graph
Even when one model is generally weaker on visual/interface work, it can still win a specific visual task if its layout choices simply match your personal reading preference better.
12Use case 11: HTML explainer doc
For a beginner-facing explainer document, more and better-annotated screenshots beat a prettier layout, because the goal was comprehension, not aesthetics.
13Use case 12: browser-use painting in Canva
Browser-controlled tasks (clicking through a real web app to complete a workflow) showed the clearest capability gap of the whole test — one model produced a usable result while the other's attempt was unusable.
14Use case 13: Skool course upload via browser use
When a task is graded almost entirely on 'did it complete correctly,' a model that fails a basic step (like a file upload) loses outright regardless of how good its formatting looks elsewhere.
15Use case 14: one-shot website clone
A one-shot creative task (clone this site's vibe) can go to the model that stayed closer to the literal reference, even if that same model is rated lower on original design work generally.
16Use case 15: YouTube 12-month review
When you tally many independent tests, the cheaper-overall model can still lose the majority of individual quality contests — aggregate cost and per-task quality are two different scoreboards, so track both before standardizing on one tool.
17Final tally and verdict
The most expensive model isn't automatically the most capable one for everything you do — testing your own actual workload beats trusting a single benchmark or a single provider's marketing.
Glossary
Terms worth knowing.
GPT-6 Astra
The AI agent tested via a Codex-style subscription in this video; asks more clarifying questions before starting and was preferred for structured, tabular, and browser-automation tasks.
Fable 5.1
The competing AI agent tested via a Claude-style subscription; tends to start working immediately without clarifying questions and was preferred for longer, detail-dense written deliverables.
Browser use
An AI agent's ability to control a live web browser directly — clicking, navigating, and filling in forms — rather than just generating text or files.
Golden dataset
A fixed, hand-verified set of example test cases used to score whether an automation or AI system is producing correct outputs when re-run.
Second brain / AIOS
The creator's personal knowledge base of notes, transcripts, and meetings that an AI agent can search and reason over to answer business questions.
Resources
Things they pointed at.
00:00toolGPT-6 Astra
00:00toolFable 5.1
37:40toolGPT 5.6 Sol
01:14toolGoogle Slides
08:15toolGoogle Sheets
29:26toolCanva
31:20productSkool
24:49toolHerk Brain (custom knowledge-graph tool)
Quotables
Lines you could clip.
00:33
“I've eaten through like three Codex subscriptions, four Cloud subscriptions, and thousands of dollars in usage credits.”
concrete, specific cost admission that builds credibility fast→ TikTok hook↗ Tweet quote
15:00
“This one was Astra, and I will say I'm very, very impressed by both of those outputs.”
the single number viewers came for→ newsletter pull-quote↗ Tweet quote
37:50
“In total, Astra was $186 cheaper, but took an hour and 43 minutes longer to run across these 15 different use cases.”
the exact cost/time tradeoff stated in one line→ newsletter pull-quote↗ Tweet quote
The Script
Word for word.
Read-along
Don't just watch it. Burn it in.
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
17px
metaphoranalogy
I just put GPT -6 Astra head to head against Fable 5 .1 across 15 different use cases that actually apply to my day to day. I'm talking things like web design, personal organization, taxes, browser use, vision, and building actual automations and software. For every single one of these use cases, I break down which model did better, how long it took, and what it cost.
So if you've been wondering what model to use for which use case, then this is the video for you. I don't want to waste any time. Let's just get straight into it today.
All right, so these are the 15 different use cases that we have. And like I said, I chose these because these are things that I actually see myself doing with these models. So there are timestamps down below if you want to jump around.
And at the very end, I'm going to show the final sort of like takeaways and final cost and time breakdown. And by the way, guys, I've eaten through like three Codex subscriptions, four Cloud subscriptions, and thousands of dollars in usage credits. So I really hope that this one's helpful.
All right, so for number one, I'm not going to read off every single prompt, but here's basically what I asked. I wanted a McKinsey level presentation on the state of SMB report for my Hercules consulting brand. So I needed this to be branded and I needed it to apply to our target audience with this fake consulting business.
So that was Claude. You can see that I gave Codex the exact same prompt. Now, immediately what I've noticed is that Codex asks more questions or GPT -6 Astra specifically asks more questions.
You can see right here, it already asked me questions, whereas Fable just took this and pretty much just ran. And that is something I noticed overall. So I just wanted to call that out to start.
So let's take a look at these two decks. Here is deck number one. I won't tell you which model it was yet, but right away, I love the way it feels.
It is capitalized. It has the branding. This feels pretty professional so far.
Now, I already did check both of these for the sources and the verification, and they are good from like... the fact perspective, right? The research they both did was good.
Now I'm looking more so at the way they actually formatted this deliverable. I think we're at a place now where GPT -6 Astra and Fable 5 .1, they're both so, so intelligent. You don't have to worry as much about fanning out agents and doing research.
It's more to me now. about how do they come back and present this in a way that tells a story and it's something that I can actually like present over, right? Because that's what we want is we want to turn that messy data into something that's readable and something that's high quality.
And right now what I'm feeling is that this is very, very branded. It matches our color scheme. We have a footer down here, which I love.
It's consistent. We have like structure. There's columns.
Like it did a good job creating this. The thing that I think right now though is it's very, very wordy as far as being able to present over something like this. it would be tough because there's so many words and it's tough to control like where people's eyes will be looking and what they'll be speaking about so as far as a presentation deck not the best but as far as like a deliverable you might turn in you might send to someone and not present this is very good and there's a lot of sources down here this came out to 35 slides now let's look at the other version which right away i don't like how this isn't capitalized i don't know i just feel like that feels very professional we have some images here and we have the color scheme as well but let's flick through we immediately notice that this one is not as wordy I immediately feel like this one's less professional though because of the structure.
The other one had like, I don't know, you could tell in each slide there was like columns. There was a little bit of like a, I don't know, there was a structure to it. Whereas this one feels a little bit more abstract.
It feels like one big element. You've got one big, you know, title thing. And I think that this one would be a much better deliverable if I was going to present in, you know, in a room full of people.
But overall, as we flick through, this one just feels less professional. The vibe I get is less Mackenzie and it's more Google Slides or like Canva. I don't know.
There's also not like a consistent footer or our logo isn't on every slide. So, you know, this sort of slide, it just doesn't feel very structured and as professional. Even the sources section is way smaller, although that does link the actual sources.
So anyways, this one was GBT6 Astra and this one was Fable 5 .1. And I will say in this specific example, I liked Fable better. So what I'm going to do here is I'm going to give Fable the win.
So I'm just going to put this down here so we know Fable won this experiment as far as just the output. But let's look at the time and cost.
Fable was 37 minutes and Astra was 23 minutes and Fable was 26 bucks and Astra was 12 bucks. So Astra was a little less than half the cost as well as, you know, 14 minutes faster. So the question would be if you used Astra and spent another, you know, 14 bucks using Astra to improve V1, would you get somewhere that you feel better about than this first version of Fable.
So anyways, right now we'll look at the time and cost at the very end, but as far as the deliverable, I liked Fable's output better here. All right, moving on to experiment number two, we did a sales letter. So basically I said, based on what you know about our certification program, I didn't give it information.
I told it to... use my AIOS, use my second brain, use any research to put together a sales landing page. So writing the copy that's optimized to convert.
So think about our avatar, think about their pain, think about all this. And I don't want it to sound AI generated. So I'm not going to read the entire sales letters.
That would be super boring. Just going to do a quick skim through and talk about which one I actually thought was better and why. So here was the one that Fable 5 .1 wrote out.
You can see it is pretty long. This one was longer. It was about 2 ,800 words.
And if we go over to GBT's version, if I come down here, it just output it right in the chat to me. This was the one that Astra wrote up, and this one was more like 1 ,700 words. So it was definitely shorter.
Now, quick disclaimer here. I am not an expert copywriter. I would want, you know...
like john or someone on my team who is really good at writing copy to tell me which one they thought is better i will say so far what i've been hearing from my team is that they are liking the gbt models lately better for writing whereas in the past claude always dominated when it came to writing but i will say if i was a prospective you know certification student and i was reading these i honestly here would have picked fable and the reason is because this is not a cheap program and Fable had things like questions people ask.
Will this get me hired? Does everyone who joins get certified? Am I technical enough?
I already run an agency. What if I miss live classes? The way it formatted this makes me feel a little bit more confident or makes me feel more aware of what I'm getting myself into.
It also had a section about tuition. It also had a section about who's teaching, who it's for, who it's not for, how this actual program works. Whereas the sales letter over here with Astra just felt a little bit more, I don't know, it just felt a little bit more high level.
And I think with the sales letter, it's supposed to be detailed enough to answer all the questions and pain that someone's... thinking about when they're considering paying for a program. So I will say in this specific example, I do think that I liked Fable 5 .1's copy more.
So here is actually what this looked like. Fable was about four minutes. Astra was about three minutes.
Fable wrote, okay, I was wrong about Astra. It wasn't 1700. It was actually more like 1300.
So about half of the words that Fable wrote. But Astra came in at $1 .43, whereas Fable came in at almost four bucks. So here, I will have to say again that I think that Fable won this challenge as far as just like...
a one -shot prompt and the initial output all right so moving on to use case number three now unfortunately guys i'm gonna have pretty much everything blurred out here because this is about my taxes right and this is one of those things where you want something that is very smart to be able to do this in a way that you trust so i'm just going to go through high level the way i feel when i look at these outputs this first one was fable 5 .1 And before I get into it, I do want to say, in this example, GBT Astra asked me like seven questions before it started, whereas Fable didn't.
So that alone makes me feel more confident in Astra's output. But let's look through here. We go through basically from January of this year to June of this year.
So Q1 and Q2 of this year. It pulls all of the numbers. It looks at federal and state.
It looks at things that I need to read. It goes to the input and it shows me all of these different things as far as the parameters, the notes, the values, the sources. And Fable gave me like a decision logic.
It said, hey, if you fit into this bucket, do this. If you fit into this bucket, do this. Whereas Astro just asked me those questions and then gave me that recommendation.
Anyways, we have a P &L. We have some projections. So for the rest of the year, what I might be estimating to pay.
We have tax calculations here for the individual return and all of the entities that are passed through. We've got payments. We've got things that are flagged.
and then we have the actual sources it pulled from so it's not bad it's a really good deliverable like light years ahead of what used to happen but now we have the one from astra we have the actual information as far as start here this is what it did we have monthly results we have the tax forecast we have okay that looks very good we have all these payments we have inputs and open items so these are other things that i might need to answer but this is just way more tailored it's easier to read as well we have the source checks we have the transaction ledger and this one is basically every single transaction here so this Oh my gosh, this is probably thousands and thousands of rows.
This is 3 ,739 rows of transactions here. And then we've got all the sources that it pulled from. Okay.
Cool. So overall, this one, I think that Astra won like no question. It was more tailored.
It asked questions. I think with taxes alone, it gave me that feeling of confidence and the structure of the deliverable was just better to me. Now, Astra did cost more here and it took longer to run.
So 40 minutes compared to 22 minutes and 22 bucks compared to 13 bucks. So even though it's more expensive, we're kind of putting that aside right now. As far as the deliverable, I liked Astra here more.
So right now we're sitting at Astra 1 Fable. two all right so this next one's interesting i asked it to go through two of my email accounts and just look through you know potentially years of emails to help me find information about subscriptions that i'm paying for how much total i've paid them you know um if prices have increased and things that i need to cancel things like that so let me pull up the two outputs all right so here is output number one from fable 5 .1 now as we get into the next tabs on the sheet.
I'm sorry guys, but I am going to have to blur some of this stuff out. It's financial information. So anyways, we have the rundown right here of the audit.
We have the different emails that looked at. We have the different types of services, memberships, programs, retainers, contractors, insurance, all this type of stuff. It's looking at the top 10.
It's looking at our different categories and it's showing us how to read this sheet. So we have all of our active subscriptions. It's showing me what email they're associated with.
It's telling me which day of the month I'm getting billed on. It's showing the current price. It's showing the monthly equivalent.
It's showing me all of the stats across all of the different payments that we've been making on some sort of recurring basis. It's flagging things and it's ranking them as far as priority. We can see the different actual flag notes.
So some of these that might be unpaid, this one's getting a price jump. This one's having a seat creep. This one's having, I have two subscriptions at once.
So maybe we need to figure out why do I have two different subscriptions on both emails there? So the point being, it's able to look through so much information and give you all of these things that are high priority and you can flag them. You can check them off.
You can, you know, delegate that out to your team. We have all charges, which actually this sheet came back completely empty. I'm not sure why, but this sheet is completely empty.
And then we have contractors and big payments. So these are things, it looks like anything above either a few thousand dollars or if it was a contractor that was like a one -time invoice, then it flagged it here as well. And now let's go over to the Astra 6 version.
You can see that they consistently just format Google Sheets a little bit different. Like you can kind of tell, Astra likes to make the cells larger and wider. and taller whereas fable usually keeps them the same and you know they still do a good job formatting but either way same sort of thing at the audit up front we have the current recurring baseline of subscriptions we have the annualized baseline tool payments tool invoices without confirmation clawed usage other bills price changes, it's flagging everything right here, priority review right here, you know, there's a discount ending, there is contract tier or contact tier, there is a price increase, there's a term change, there's seats and tool overlap, so it's flagging all that right away.
We then have all the subscriptions, so this shows me a list of every single one, it gives me the price, it gives me the last known bill. Now this isn't showing me what... date oh there it is yeah it's showing me every single day of the month that we have the next one we have other bills as well so these are kind of the bigger like bigger contracts um contractors or or bigger bills we have payments so this is also like every single one and i think this is what fable was trying to do on this tab with all charges but for some reason just didn't come through so we have this here And then we have also other notes as far as different topics and different things to sort of be aware of.
So this is kind of like what it was flagging. So overall here, I will say that Astra gave me a better output. It just, you know, Fable messed up one sheet completely.
And Astra was cheaper here. So 12 bucks, 12 and a half bucks compared to about $50, which is insane. $50 for that output from Fable, crazy.
It did take 17 minutes, whereas Astra took 30 minutes. But anyways, right now we are sitting at two to two, Astra and Fable. All right, moving on to number five, we have meeting analysis.
So I asked it to look through all of my leadership meetings and my syncs with John and looking at other meetings that are relevant to the team. And I wanted it to understand, or I wanted it to tell me what are the biggest pain points, three biggest pain points, and one of the highest value automations that we could build to remove some constraints and help us scale faster.
So Fable gave me the biggest pain points of me being the single production and review engine, which is something that I've talked about 25 of 58 meetings. Certification delivery does not scale past one cohort. So we're trying to figure out what that will look like.
The critical constraint of the business is also a large cost center, 52 students, blah, blah, blah. And then we also have demand collapsed and every fix is blocked by unknown work and untrusted data. Weekly leads fell, blah, blah, blah.
Yeah, all of this makes sense as far as our three biggest pain points. And then the one automation it came up with was an AI customer success layer for the cert. One live record per student built from circle activity, mastery check outputs and attendance.
Students upload case readouts to an endpoint instead of DMing Pat. Weekly due dates and reminders to non -submitters go out through Gmail or Beehive. I think this is something that I've wanted to do and that we've talked about a lot.
That makes a lot of sense based on all these conversations. And I also wanted to note here that Fable 5 .1 analyzed 58 meetings. And if we go over to Astra, Astra analyzed 79 meetings.
So it did get the context of 21 more meetings in here. But anyways, let's take a look. The first one was that growth depends too heavily on a few acquisition channels.
Yep, same pain point, I think, as Fable pointed out. The second one is that the customer journey requires too much personal intervention. This appeared in an undefined AIS Plus onboarding call, a resurface during certification.
And leadership still has to reconstruct status and clarify who owns what. Okay. interesting.
So these are different pain points than Fable 5 .1 found. Now, as far as the automation I would prioritize, it came up with the same thing, a customer success workflow, which would check each student's onboarding status, unanswered questions, available progress, identify missteps, prepare a follow -up, all makes sense. Now, in this case, I think Astra won again.
Like before I even thought about the cost, which is absolutely absurd, Astra still won because I liked its recommendation about the leadership stuff and ownership. And I think that's something that you really want to... make sure you have figured out within a team.
But anyways, if you look here, Astra was five and a half bucks and Fable 5 .1 here was 46 bucks, which just makes no sense to me. I'm very, very confused on how it costs that much. I had to look at the session logs.
I also did a slash usage and this one cost 46 bucks, even though the active... agent time was just six minutes so i'm very confused on where those tokens went like i don't understand why it also analyzed less meanings in astra but anyways that's pretty interesting this winner is definitely astra so now we are sitting at three to two astra to fable All right, so moving on to number six, this one's super interesting to me.
I gave it this folder, which has over 150 gigabytes of resources and recordings from our AIS Live event in July. And I asked them to make us basically a recap, like an energetic, real 60 seconds that shows like a recap of the event that we can use to promote the next ones and things like that. Now, I'm really impressed with both of these outputs, let me just say, because with how much information it had to look through and...
how it had to create a story and figure out how to put everything together with music. I think they both did really, really well. So here is Fable 5 .1.
Hello, hello, AIS Live. Let's get some energy going in here. Oh my God, I'm so excited for this.
You guys are throwing a fantastic event. Literally like logged off yesterday, just buzzing. I had goosebumps up.
Let's get you guys paid.
This was an experience that didn't feel like a webinar.
Okay, so that was Fable 5 .1. Let's take a look at Astra's version. This has been so energizing.
Yesterday, I did a workshop where I went all in on every little tool that I use.
Plot code is my primary driver. And then I built like my entire second brain slash personal agent on top of it. I could go to bed, come back, and it'll still be working, you know, 10 hours later to try to achieve that goal.
It's great to be here. You guys have done a fantastic event. Everything's been super smooth.
You guys have been such an amazing audience today.
Now this one, like I said, is really tough. I think that they both did different things well. I think that Asher did a better job of putting like live clips in there.
I think that Fable did a better job of kind of like gearing it up for the next event in October. There were a few moments where there were some things that were off. Like Fable had a few moments where the timing and the voices felt like not very synced.
Asher had one weird moment where it was showing a video of Russ, but it said Devin Kearns. And it was just like, that was clearly off. But I am going to give the slight edge here to Asher, but I will say I'm...
very, very impressed by both of those outputs. I thought it was given the vague goal. They both did a really good job of exploring and creating a story.
But anyways, this one is going to go to Astra because you can see the stuff here. Astra took 30 minutes, whereas Fable took 50 and Astra cost 16 bucks and Fable cost 26 bucks here. Now, next we did sizzle reels.
I basically gave it an example of a sizzle reel that I really liked. It was sort of like a SAS, I don't know, a motion graphic, a motion design video that I liked. And I told it to analyze what was good about it.
And I told it to use that as inspiration and create us one for AIS. So let me show you guys these outputs. First one is Fable 5 .1.
Okay, now let me play you Astra's output.
Okay, so Astro wins this one. I mean, I will say that music is a little bit annoying. But besides that, it took screenshots from AIS.
Like it went in here and it grabs pictures of the actual courses. Where was the other one? It took a picture here, which I thought was just like, I love that idea that it's thinking, okay, why don't I actually go get real proof and screenshots from the thing I'm actually building a reel about and put it in there.
I also love this, like these 3D designs. I don't know. It just felt like it had depth.
You can see the shadows here as it kind of. you know, zooms in. I thought that this, as far as the energy that it had was much better and much higher energy than the first one from Fable 5 .1.
Now let's take a look at the cost and time. This one was about similar time, but Fable 5 .1 was, you know, a little over half. So 16 bucks compared to 26 bucks.
But once again, I still think that in this case, I'm going to take that output from Astra as the winner. So moving on to number eight, this one's a little more out there as far as I probably wouldn't really be doing this every day for work, but I had them build me these games. So we have two different games where you're a pistachio and you're trying to escape the kitchen.
So I'm just going to go through these real quick and you're going to tell me which one you thought was, you know, sort of like which model. So you can see here, it's pretty smooth. I can double jump, I'm collecting these like cereal bits and I have to hit these checkpoints to get to different spots in order to get out of the kitchen.
The physics are real, like I'm bumping into these things and I can like, you know, hit the spoon and everything like that. It's pretty easy, but you know, I missed a cereal, but I'm able to escape and win the game in about 43 seconds. Okay, now here is the other version.
This one definitely looks a little bit more, I don't know, it looks a little bit more premium, doesn't it? This obviously, this wall isn't there, but it's just because we want to be able to see like this. But this one feels a bit more premium, I'd say.
It also feels smoother when I'm actually controlling this. It is guiding me very specifically, though. I don't know.
I can double jump, I think. Oh, no, I can't double jump. The physics seem to be real still.
I can bump into stuff. I can move things. I can fall into there.
We've got a hot stove. Oops. It's interesting, though.
They both decided to have, like... The same flow as far as you jump up on the drawer and then you jump up on books to start. You're now on the stove.
You're going over to the sink. I mean, that's how most kitchens are. But still, the actual flow is pretty similar.
And now we escaped in this one in 46 seconds. So similar amount of time. I did get all of the crumbs in this one.
They were designed very similarly, but I will say I think that as far as like feeling realistic and, you know, physics, I liked this version better. And this version was Astra. Anyways, look at these stats.
They were very, very similar. So Astra was eight bucks cheaper, nine bucks cheaper, and it also took about 40 bucks longer. So very similar stats.
Nothing blew anything out of the water here, but I would say that I'm giving this one to Astra. Once again, that one's subjective. Maybe some of you guys liked Fable 5 .1 more, but what I think is really important is that you're applying your own taste and you're finding your own use cases because some of these use cases, you guys might be like, well, I actually liked Fable 5 .1 here.
Cool. Then you just figured out maybe you use Fable 5 .1 for that use case, even though some other people might use Astra for that use case. It's all about your own use cases and your own experience.
experiments. Okay. So moving on to number nine, we have an actual SAS, like kind of a mini SAS, and it's an evaluation app.
So you should be able to put in like automations and then run them against some golden data set and see the score and understand like what you need to fix. And then you can rerun them and you can generate reports and things like that. So you can see here, this is the first version.
This is Fable 5 .1. I think that right away, this is. a little bit more of an intimidating interface it's very very wordy but we can see all of these tests we can see our different automations here on the left as well we can add a new one so if we wanted to put the name a description and then actually you know say how do you actually trigger this automation is it python or do you hit an api and then from there let's just put in some information you would actually put in the url or you would put in the python script you would put in the test cases and then you would write you know how to grade it so that is going to be very valuable.
And it also has all of this stored on the backend because it has to store all of these actual runs and it has to store the code and it has to show the improvements and things like that. So here we can see this support ticket one. I think they ran two tests.
You can see that there's also like a visual breakdown of what this actually does. So this is a very deterministic script. We can see the code, we can see the setup, we can see the test cases.
So this is essentially our golden data set. There's only 12 examples. You probably need more than that.
But this was all... generated to actually test the system, right? And if I actually go in here and hit run, what does this do?
You put the name, you put what changed, you put any tags, and you compare it with a different version if you want, and then I just go ahead and start the test. And so this came back, you can see, I mean, there were no changes, so we still got two out of the 10, or sorry, two out of the 12 didn't hit right. It also went super, super fast, because like I said, this is super simple, but then you can generate a report.
So if I come in here, you can see all of the different... tests that we've ran and it would show us like what's still broken. It would show us the improvements we make.
So this is something that, you know, I actually do want to keep building out so we can do this for, you know, our students and basically give them a better way to evaluate their agents and things like that when they're building them out because evals, obviously we all know are super, super important. And it's also good to be able to see them visually.
So this was Fable 5 .1's version. Let's look at Astra's version. So I do like the interface of this one more as from a design perspective, I think that this one feels, I don't know, it just feels a little bit more premium.
It's also less wordy, so it's less intimidating right away, I think. We have a test library, so we can see all of these different versions. We have run history.
We also have reports over here that we can look at, so we can see customer support agent. I like how this report is structured. Let's see if we look at this next one.
We can see how it changes over time. This one feels more branded. This feels like something I'd rather get in front of some eyes and get some feedback on compared to, oops, I just closed out of the other one, compared to this version.
This one just feels... way more like a POC compared to this, I feel like is something we can almost get out there and start testing against. So if I come into an automation real quick, let's say we go to the support ticket router and I want to go ahead and see it visually.
I like that a lot. We can see all the previous runs and I can go ahead and run this eval again and we can see it running a bit better. We're seeing it actually go through the test cases.
I like this experience way, way more. And look at this. Astra came back and cost half the price of what Fable spent in order to do this.
It did take a little longer, two hours, 25 minutes compared to an hour and 20 minutes, but it's just so much better for less cost. So in this case, I'm giving this to Astra once again. So Astra is now up, I believe seven to two on Fable.
All right, moving on to number 10, we now have this Herc brain visual. So this is a local host where I can see. basically my brain.
We see different clusters over here. This seems more like YouTube videos. This seems like things that are actually going on in the business.
Oh wait, we can see down here, cloud memory, video knowledge, and we can remove things or put them back in. Business wiki, skills and agents, projects, and this is dynamic. It's really cool because we can see only what needs attention.
We can see show all labels. We can see 612 notes, 2 ,500 connections. We can analyze things.
We can talk to it. We can search for things. Let's see if I search for personal.
OS, personal assistant, open claw. We can open up the specific things and that is a pretty cool visual. So we can actually start to see and hear inside of the brain.
And as we click on different things, we see all of the connections that are linking into it. So all the relationships. So that is a pretty cool experience.
It does feel pretty smooth and I can control this, eh, not too well. Like I feel like I have to click on something to make that the center, but I think that that's pretty cool. So that was Fable 5 .1's version.
Let me now open up Astra's version. Okay, so this one's a little bit different. I can still click on all of these things.
These nodes are smaller. I don't like how that feels. I don't like this interface at all, like nearly as much.
I have the same sort of vibe. I can click on business strategy or AI and ideas, videos and content, people, projects, book and writing, all knowledge. This interface is more overwhelming for sure.
Like I definitely like this one more. I wanted just something simple where I can see the brain a bit more visually. And I was honestly expecting Astra to come in here and kick Fable's butt here, but I think that I like Fable's more.
And now let's look at the cost. Astra took 40 minutes, Fable took 24. Astra and Fable basically cost the same.
Fable was a little bit cheaper, but I am gonna give this one to Fable. I think that Fable gave me more of what I was looking for. You know, once again, kind of a subjective grading criteria, but I think Fable won here.
So what does that put us at now? We're at seven to three, Astra to Fable. And we have five more use cases left.
So Fable could still take the lead. All right, guys, so this one is an HTML explainer doc. What I did is I gave it my YouTube video, the 25 Grokbot concepts explained, and I told it to explain it to me simply, simply, simply, and visually, and I told it to take screenshots.
So I said, basically, give me a doc that I can read through, and it feels like I actually watched the entire tutorial. So we have all these different parts. We can see the screenshot to start.
This one obviously is kind of wordy up front, but that's okay. We see how the team fits together. So me, these executives, these operators, we can see all these different plugins and computers and skills.
So that's a good little explanation at first. Now we move into the concept. So we're taking screenshots and it is being analyzed.
This is Fable 5 .1, by the way. We can see the executive bots. We can see the name and job label.
These so far are coming through well as far as these screenshots being cropped and annotated, which is really nice. It's taking a lot of different screenshots. We have the buttons here, we have all this.
I think that this is coming out pretty well and it's showing us what we need to see from the actual video. Now I do think this is pretty wordy. I think that, but it is getting very specific.
Like I do think I could give this doc to some people and they would be able to feel like they actually watched this video. It's cropping things, it's zooming in, it's annotating. I think that this is doing a really good job.
It's even showing all of this. Yeah, I mean, this is doing a really good job here. So let's go over to Astro's output.
This one already, it looks a bit better as far as the HTML. And now let's see if the screenshots are good as well. So we have one here.
It's pointing at a bot. It is pointing at the description, the executives. It's pointing at these different buttons.
I will say Fables, it was more detailed for sure. Like this one was designed better. But Fable's screenshots were better.
Like Fable took more screenshots and made more annotations. And so far, I think that that was a better learning experience. Even though this one, like I said, looks a little bit more, I don't know, it feels prettier or better.
But if you were really thinking about, could you hand this doc to a complete beginner and have them actually go through and feel like they're learning? I think that Fable is going to take this one. But let's look at the cost.
Like I still am going to give the output to Fable. But when we look at the cost. Astra was much cheaper here.
So the question is like, you know, if you used Astra and had it spend another, you know, 20 bucks or so, would it be better than Fable's output right here? And I would argue that yes, it would be better. So it was way more efficient here, but it's, you know, it's V1 wasn't as good as Fable's V1 in my opinion.
So Fable's going to take this one. So now we are at seven to four Fable, seven to four Astra to Fable. Okay.
So moving on to number 12, we have this sort of like painting example in Canva, which is a bit silly. but it does highlight the browser use and the vision really well because GBT6 Astra had to look at this picture, open up Canva, understand how to navigate with the different tools and things and build this out, which it did.
And I think that this does look pretty good, especially when we go over to Fable real quick and I open up Fable's output. If I go to the painting and we open up the browser here, this is literally what it gave me. And this was the reference image.
I mean, this is just, this is laughable. This is laughable. I was expecting Fable 5 .1 to do something much better, but it like, I don't know what exactly it tried to do, but this is obviously just very bad.
So no question, Astra wins this one, and Astra was also cheaper. So it's pretty ridiculous here how much Fable spent and how long it worked just to give me that crazy output. So Astra wins.
Okay, so moving on to the next example, number 13. What I did is I gave... Fable and Astra, this doc, which said, hey, this is a new course.
I want you to put it into the free community as a draft. And I want you to use browser use in order to do this. So they basically opened up school.
They went to the classroom. And they put in these draft courses, Astra AIOS and Fable AIOS. And they put in their different courses.
So they were basically instructed to have the description, the key points, the key quotes, and any resources. And they had to put the video in here. And so this is Astra's output.
I told them to do just the first five. And this looks really good. It had to navigate the interface.
It had to basically come in here. make a new page, it had to put in the video, it had to put in the text, and it did it really well and saved everything here as a draft. Now if we go to Fable's output, there are no videos here.
This one arguably might have been structured better as far as like this looks cool and these are numbered, whereas the other one with Astra, it wasn't like numbered as well, like it just, it's longer. But the big, big issue here is that Fable ran into an issue and for some reason couldn't... upload the video.
There was some issue with having it downloaded locally and whatnot. You know, I gave them everything it needed. I gave them the videos.
I gave it the links. This is something that Asha did way better. And I will say I've done this multiple times in my school and I've used 5 .6 Sol with browser use and just in general.
Codex and the GBT models are significantly better with browser use. So this one I'm going to give to Astra. And honestly, the cost and the time wasn't too different here.
Very similar and very similar. Astra was a little bit cheaper and Astra's results were much, much better. So this one's going to go to Astra as well.
So we are now at the place where the score is Astra has nine and Fable has... four. So Fable can no longer catch up, but let's look at the last two outputs as well.
Okay. So for number 14, we did a website and basically what I wanted to happen is I gave it this URL and I said, Hey, I want you to basically help me clone this website. I like the structure.
I like the feel. And I want you to make this for perk form, which is the fake, you know, um, can company. So this is a really cool scroll dynamic.
I know you guys, some of you guys hate this sort of site. That's not the point. The point is could Astra and Fable analyze this site.
and make one that has the same feel and the same vibe for perk form. So here is Fable 5 .1's version. It loads up.
We have this, you know, 3D dynamic element, coffee reformed. We have like some, you know, DNA looking stuff here. We keep scrolling down.
We've got these coffee beans falling. We've got the can come back. We've got all of this going on.
I would say that it's doing a decent job here. It's even got the huge like zoom out. It's doing a decent job of making this feel like the other site.
You know, obviously there are some issues here, but I think it's doing a decent job. Like for a one -shot prompt, although sometimes I get into these issues with the scroll. I mean, that's a bit of a bug, right?
That's when you hit the bottom of the site, something weird goes on there. So there's obviously some bugs, but it's not terrible for a one -shot. And now let's take a look at Astra's version.
We have the big can here, which is a little bit dynamic. There's something going on with this text, right? Like the coffee's out of bounds for some reason.
Keep scrolling down. I do like this animation. I mean, that's pretty cool.
It's a coffee bean. There's definitely layers in the back. And by the way.
This one, I didn't tell it to use any skills. I just said, hey, I want the same vibe. I want you to analyze the original site, this one right here, and make me one like this, right?
So anyways, we're coming through the little sort of like DNA thing, coffee and protein together, two drinks in one can. I like that animation. This one overall does feel smoother than Fables.
I do think that as far as a first pass, I like the way this one feels. We get stuck at the bottom and we'd have to scroll. all the way back up to the top now.
I do think that Fable did a better job of actually just recreating what we saw in the original site, which was the prompt. Now I do get stuck. There's a little bit of a, there's bugs like this.
Like I get stuck, but as far as recreating this, Fable's felt more similar to this, I think. So we will give this one to Fable. It kind of needs it as well, but I still think in general, when it comes to design, you guys have probably seen throughout this video and you've seen my other videos.
In general, I do think I like Astrum more for design, but that was like. hey, show me what you can do, take this URL, build me something crazy in one shot. It was kind of a weird prompt, right?
But in this case, I still think that Fable won, but here's a look at the cost and the time for this example. All right, and this is the last one, number 15. We have YouTube 12 -month review and strategy.
So it basically had to look through the past year of my YouTube stuff and give me some stats. So this is interesting. I mean, these are big docs.
The one on the left is Fable, it's 25 pages. The one on the right is Astra, it's 15 pages. So let's just slowly scroll through and see what's happening here.
astro right away has you know this chart on new viewers and regular viewers whereas over here we have you know, the year in numbers. We have a monthly review.
We have total views across the past. Oh, interesting. So actually what they show me first on the left side is from January to now, whereas over here I'm getting sort of like the actual past 12 months.
So those stats are going to be a little bit different. It's showing me content type. It's showing me the monthly scorecard.
It's showing what my best content actually does. It's showing me some packaging stuff. It's showing me what to worry about.
So it says news brings the views, but does not bring the subscribers. It's saying the news packaging is drifting towards hype. The audience is saying so.
It's saying. Agency mechanics content has stopped working on YouTube. That definitely feels to be a bit true for this audience.
Now it's deep diving into some packaging. So it's analyzing my titles over here. Whereas over here, when we have the packaging, what I would keep and improve, the evidence, retention, first 30 seconds, who you are reaching and how.
So it's looking at the age, it's looking at the gender, it's looking at the geography. It analyzes what the comments are asking me to do. So then at the end, they kind of like put all this stuff together and they give me the next 90 days and they give me some things to look at overall.
they're very similar i do like the way that this was packaged more by astra it just feels a little bit easier to read and it also feels more visual so i'm going to give this one to astra as well and you can see here though asha did spend more and take longer so 42 minutes compared to 17 and 27 bucks ish compared to nine bucks so astra wins but it was longer or slower and more expensive here.
But anyways, guys, that was 15. That was 15 different use cases here that we scored and looked at the cost and time. Astra won 10 of them and Fable won five of them.
But let's take a look at the total numbers here. I had both Astra and Fable give me these stats to make sure they were consistent. And you can see the numbers are exactly the same.
Fable was nine hours, 35 minutes and 45 seconds of total runtime here, costing us 513 bucks, 36 cents. And Astra was 11 hours, 19 minutes, 24 seconds, and cost us 326 bucks, 98 cents. So in total, Astra was $186 cheaper, but took an hour and 43 minutes longer to run across these 15 different use cases.
Now I will say this was 15 use cases. In general, my experience has been that Astra feels quicker and feels more efficient. And also it doesn't give us that five hour window in Codex.
And it doesn't give us that fable limit. You know, we can just use Astra the whole week. So I do like that.
We also get more inference with Astra in general because of the codex subscription. So right now I'm. heavily leaning towards Astra for my day -to -day, my use cases.
But the other thing I want you guys to keep in mind is that GPT 5 .6 Sol is so, so good still. It's so capable. And for the majority of my knowledge work, I could use that just fine.
And that even honestly might be overkill for the majority of stuff that I'm doing. So it's not like you need to use Astra only for all your use cases. And it's not like you need to go grab Fable 5 .1 for all your use cases.
I'm still a very firm believer in they both have different strengths and weaknesses. They both have value. And that's why I love doing things like this to find out what are the different use cases and how do they, you know, know, behave in different scenarios.
And so, yes, I am going to be using Codex more right now because that's where it currently is, but this stuff shifts so fast and I will always keep both subscriptions or multiple of both subscriptions so I can keep testing them out as things change, as my use cases change, who knows what might happen. But anyways, that is going to do it for this one.
I hope you guys enjoyed this breakdown. I hope that it was insightful or you learned something new. And if you did, please give it a like, it helps me out a ton.
And as always, I appreciate you guys making it to the end of the video and I'll see you guys all in the next one. Thanks everyone.
The Hook
The bait, then the rug-pull.
Fifteen real work tasks, two frontier AI agents, and the exact cost and time for every single run: the creator burns through three Codex subscriptions and four Claude subscriptions to find out which model actually deserves the day-to-day seat.
Frameworks
Named ideas worth stealing.
00:43list
The 15-use-case agent scorecard
Research / branded deck
Sales letter copy
Taxes / P&L
Subscription audit
Meeting analysis
Event recap video
Sizzle reel
3D game build
Eval SaaS build
Knowledge-graph visualization
HTML explainer doc
Browser-use painting
Browser-use course upload
One-shot website clone
YouTube channel review
The creator's method for comparing two AI agents: run each through the same 15 real, self-sourced tasks and score output quality, elapsed time, and dollar cost independently for each one, rather than relying on a single benchmark.
Steal forAny evaluation of two AI tools/subscriptions where the deciding factor should be your own workload, not a published benchmark
CTA Breakdown
How they asked for the click.
VERBAL ASK
38:15subscribe
“if you did [enjoy this], please give it a like, it helps me out a ton”
A single soft ask at the very end after the full breakdown, no hard sell or mid-roll interruption.
FROM THE DESCRIPTION
PRIMARY CTAWhere the creator wants you to go next.
Add Modern Creator as a preferred source and Google shows you more of our breakdowns in Search, Top Stories, and AI Overviews. It only changes what you see, and you can undo it in your Google settings anytime.
Add to Preferred SourcesOpens your Google source preferences with us pre-loaded. Tick the box and you're done.
A hands-on tour of a synced, phone-controllable AI agent team — agent computers, teachable skills, scheduled routines, event triggers, and where it stops making sense versus Claude Code or Codex.