Claude Code vs. Codex: An 8-Site Head-to-Head Design Test
Same prompts, same brand kits, two coding agents — one landing page built by each, eight times over, scored on design and on the bill.
August 26thA screen-recorded walkthrough of turning a scroll-captured reference video into a fully rebranded landing page, switching between GPT-5.6's Sol and Terra models inside Codex as the design gets refined.
A single AI prompt only produces a starting point, not a finished design, and the real skill is matching each editing task — full structural rebuilds versus small fixes — to the right-sized model.
The video tests OpenAI's new GPT-5.6 model family — Sol, Terra, and Luna — by rebuilding a landing page inside Codex. Instead of screenshotting a reference site, the creator records a slow scroll-through video so the AI can see load animations, transitions, and scroll behavior, then prompts Codex with an explicit keep-this/change-this split: preserve structure and motion, replace brand, copy, and color. Sol handles the full video-analysis rebuild; Terra handles the smaller fixes after — a broken card-drag loop, a flat pricing section, a weak hero. The core conclusion: one-shot prompts are always a starting point, never a finished product, and picking the right model per task matters more than always reaching for the most powerful one.
Sign in and you get 23 free chat messages on us — ask for the hook, quote a framework, find the exact transcript moment, generate a markdown action plan. Bring your own key when you want unlimited.
Create a free account →
Introduces Sol, Terra, and Luna, and narrows the video's focus to Sol and Terra — the two models most people will actually use inside Codex.

Argues the final result matters more than a single flashy one-shot prompt, and that a strong prompt still requires writing skill and experience.

Picks a random inspiration site on Framer, then records a slow scroll capturing unload animations, layout, scroll effects, and transitions rather than taking a screenshot.

Uploads the scroll-capture video to Codex and writes a prompt with an explicit keep-list (structure, layout, animation, interaction) and change-list (brand, copy, images, color).

Reviews the first Sol output — new AI-generated imagery, a hero section that needs stronger word reveal, and a section meant to be sticky that isn't pinning yet.

Calls the first pass an impressive but imperfect starting point, then brings in a second animation reference to combine ideas rather than clone one source.

Explains why the full-video, multi-section combine pass stays on Sol, and states plainly that anyone claiming a true one-shot result is describing a project they'd already finished before.

Walks the updated page and flags the pricing and FAQ sections as still boring, while noting the testimonial section is strong but missing interaction.

Moves from Sol to Terra for everyday iteration — quick design changes and smaller improvements rather than another full structural pass.

Diagnoses a broken infinite-loop drag interaction on the pricing cards, writes a precise fix prompt for Terra, and switches to a cloned AI voice mid-recording after losing his voice.

Calls the hero section the part that normally takes the most time, points to animated imagery and tools like Spline or WebGL as options, then signs off.
One AI prompt only produces a starting point; the real work is a keep/change brief up front and matching each fix afterward to a model sized for that task.
“Let's be honest, creating a very detailed prompt that give you a beautiful blended page in just one shot still require AI strong writing skill and a lot of experience.”
“When people say, I just made it in one shot pro, that's a lie.”
“The only way to have a one shot prompt is when I'm done with my landing page — I just create a sort of one prop shot with the code inside the prompt so you will get the exactly same result.”
“The goal is not to copy the original website — at the end, the result must be different.”
See every word as it's spoken — crank it to 2× and still catch all of it. The same dual-channel trick behind Amazon's Kindle + Audible.
OpenAI just shipped three GPT-5.6 models instead of one, and the video's real question isn't which is "best" — it's which one earns its place at each stage of actually building something.
The prompt structure used to hand a reference video to Codex — explicitly separating what the rebuild must preserve from what it must replace, so the result reproduces the motion system without copying the brand.
The video's working rule for which GPT-5.6 model to reach for at each stage of a build, based on how much reasoning and cross-section context the task requires.
00:00
00:11
00:19
00:27
00:35
00:42
00:50
00:55
01:09
01:10
01:21
01:29
01:37
01:46
01:53
01:57
02:08
02:16
02:24
02:32
02:39
02:47
02:55
03:03
03:11
03:18
03:26
03:34
03:41
03:46
03:57
04:05
04:13
04:21
04:29
04:36
04:47
04:48
05:00
05:08
05:14
05:25
05:31
05:39
05:44
05:54
06:02
06:10
06:18
06:26
06:33
06:41
06:49
06:57
07:05
07:12
07:20
07:28
07:36
07:44
07:51
07:59
08:04
08:15
08:23
08:30
08:38
08:46
08:54
09:02
09:09
09:17
09:23
09:33
09:41
09:48
09:56
10:04
10:12
10:20Add Modern Creator as a preferred source and Google shows you more of our breakdowns in Search, Top Stories, and AI Overviews. It only changes what you see, and you can undo it in your Google settings anytime.
Same prompts, same brand kits, two coding agents — one landing page built by each, eight times over, scored on design and on the bill.
August 26thA designer builds a construction company's landing page in Figma, generates every photo and build-timeline video with AI, then hands the whole thing to Claude Code to assemble.
August 9thRiley Brown and Ras Mic dig into GPT-5.6, Codex's background computer-use, and why self-scoring agent loops are turning coding tools into a general operating system.
July 12thOpenAI turns ChatGPT from an answer engine into an agentic coworker — a new Work mode, a desktop app that operates your files and apps, and shareable AI-built websites, all riding on the GPT-5.6 model family.
July 9thTheo spends a week testing two rival "skills" repos for AI coding agents, Matt Pocock's 215,000-star collection and Cursor engineer Lauren's PStack, and finds the real value in a handful of specific files, not the whole install.
August 19thA free Claude Code skill interviews you about your brand and journey, generates or reuses your assets, then builds a scroll-synced site and checks its own work before handing it back.
August 22nd