GPT-6 Astra's Computer Use Is Ridiculously Good
Five ways Mark Kashef points Codex's computer use at apps that have no API, from flight searches to phone settings.
September 5thA 10-minute live walkthrough of five agentic scenarios where /goal cleans, sharpens, revives, forges, and maintains your Claude Code OS by itself.
You can use /goal to automatically optimize, audit, and maintain your agentic OS by delegating non-technical cleanup tasks like skill consolidation, rule contradiction detection, dormant project revival, and scheduled maintenance to Claude.
The /goal slash command in Claude Code and Codex runs an agent in a loop with a separate judge model verifying completion, and you can point it at your own agentic operating system instead of formulaic build tasks. The mechanism is simple: write an objective under 4,000 characters, optionally pair it with a rubric file for scoring criteria, and let the dual-agent loop iterate until the judge confirms terminal state. Five practical applications cover the lifecycle: clean bloated skill folders down to essentials, sharpen individual skills against your own rubric, revive dormant half-built projects, forge new skills by mining recurring patterns from session transcripts, and maintain the system on autopilot by chaining /loop with /goal so cleanup runs every thirty minutes against a maintenance log.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Create a free account →
Reframes /goal as an OS maintenance tool. Lists the three familiar agentic OS problems: bloated skills, expanding CLAUDE.md, contradicting rules.

Goal input (<=4000 chars), primary worker agent, judge on a different model checks terminal state. Excalidraw WORKER to JUDGE to DONE? diagram shown.

47 skills + 7 rules in a sandbox folder. /goal audits, consolidates, archives. Result: 47 to 17 skills, 7 to 4 rules, 3 contradictions resolved. 2m 50s.

Early adopters community pitch.

rubric.md with 5 criteria. /goal simulates thumbnail-prompt-builder, scores vs rubric, rewrites SKILL.md, logs to ITERATION_LOG.md. Key: write your own rubric or AI grades itself easy.

22 projects folder. /goal checks git log, attempts to run each, fixes deps, rewrites README, logs verdict to ALIVE.md. 18 stubs archived, 4 kept. 2m 16s.

50 JSONL session transcripts. /goal finds 3 recurring patterns with no skill: Excalidraw 8x, content audit 7x, LinkedIn 6x. Creates all 3 SKILL.md + SMOKE_TEST.md files.

/loop + /goal combo. Three autonomous trigger types explained. Live demo: cron created, drifted-skill archived, contradiction resolved, MAINTENANCE_LOG.md written, loop complete on turn 1.

Prompts in description link 2. Community link 1. Claude Code Zero-to-Hero course upsell.
/goal is not a task runner — it is a quality loop you can point at your own system.
“We're giving the agent a mirror. We're telling it to go through all of the assets to see how it could best optimize itself. The target of the goal is optimizing the very system trying to achieve the goal.”
“It will go easier on itself to try to increase the chances that it accomplishes the goal.”
“Way easier than having to remind yourself to do it every day.”
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
Everyone has seen the build-me-a-snake-game demo. Mark Kashef opens by asking a different question: what if you pointed /goal at the system doing the building? Five live demos later, your skills folder, CLAUDE.md, and rules files have reorganised themselves — and you did not touch a single file.
Two different models in a loop: one does the work, one audits it. Runs until judge says condition is met.
Complete taxonomy of what /goal can do to your agentic OS. Each mode has a distinct objective shape and a different folder as the target.
Three ways Claude Code runs without you. Can be combined: /loop 30m /goal <prompt> runs a goal on a cron schedule.
If you let the AI define success criteria, it sets an easy bar. Define rubric.md yourself with specific pass/fail criteria per criterion. Goal only clears when YOUR criteria are met.
“If you wanted access to the prompts that I showed you so you could use them or take derivatives of them for your own use case, I'll make them available to you in the second link in the description below.”
Dual CTA: prompts in description (high intent, immediate value) + community link (recurring engagement). Course upsell shown visually without hard pitch. Well-executed.
00:00
00:08
00:19
00:28
00:39
00:40
00:51
00:59
01:07
01:12
01:22
01:30
01:38
01:45
01:53
02:01
02:07
02:12
02:16
02:24
02:31
02:39
02:48
02:56
03:03
03:10
03:18
03:29
03:37
03:45
03:53
04:01
04:09
04:17
04:25
04:33
04:41
04:49
04:57
05:05
05:14
05:22
05:31
05:40
05:48
05:57
06:06
06:14
06:23
06:32
06:40
06:49
06:58
07:06
07:15
07:24
07:32
07:41
07:52
07:58
08:06
08:14
08:19
08:30
08:38
08:47
08:56
09:05
09:14
09:23
09:33
09:42
09:51
10:00
10:09
10:18
10:27
10:35
10:40
10:46Add Modern Creator as a preferred source and Google shows you more of our breakdowns in Search, Top Stories, and AI Overviews. It only changes what you see, and you can undo it in your Google settings anytime.
Five ways Mark Kashef points Codex's computer use at apps that have no API, from flight searches to phone settings.
September 5thGrok Bot, Hermes Agent, Claude Cowork, ChatGPT Work, and OpenClaw are all built from the same six parts. The video ranks them twice, then argues the strongest long-term move is skipping all five.
August 23rdA creator walks through the local tool he built that routes to 37 image and video models, then reverse-engineers exactly what a subscription platform like Higgsfield is doing so you can rebuild it yourself.
August 15thA breakdown of why maxing out effort settings on Claude, GPT, Grok, and Gemini rarely makes the output better — and the framework for picking the right level every time.
July 13thA 9-minute screen demo of /power-up and /insights, two Claude Code slash commands that most users have never touched.
April 3rdA 24-minute practical walkthrough of the 15 features Boris Cherny (Claude Code creator) flagged in his 2M-view tips thread.
March 31st