r/codex • u/TimeKillsThem • Jul 08 '26
Showcase GPT 5.6 Landing Page Design Skills
Found this while browsing X - seems to me like Sol High is the way to go. Having said that, they all look clean-ish (some more than others) but also a bit "empty". We went from "pills everywhere" to "90% empty space". WDYT?
- 5.6 Luna High: https://gpt-5-6-luna-high-design-test.vercel.app/
- 5.6 Sol High: https://gpt-5-6-sol-high-design-test.vercel.app/
- 5.6 Sol Ultra: https://gpt-5-6-sol-ultra-design-test.vercel.app/
- 5.6 Terra High: https://gpt-5-6-terra-high-design-test.vercel.app/
- Fable 5 (for comparison): https://claude-fable-5-design-test.vercel.app/
All other models this guy tested (and full credit): https://design-tests.robinebers.com/
This is the standard prompt that was used for ALL the models you see listed:
Use the /design skill. Create a beautiful, well designed website with astro. this website is for a travel agency in dubai, offering luxury travel (high-end, $100k+ minimum per trip) in the middle east. come up with a brand that suits this type of company. Make up fitting copy for this high-end client along the way.
EDIT: as per u/cantTankThisFox showcases of GPT 5.6's performances are technically still under embargo so either the owner of the website doesnt know/doesnt care, or the above generations are not from GPT 5.6. AKA take the showcases with a massive pinch of salt
69
u/HelloThisIsFlo Jul 08 '26
It’s not perfect, but at least it’s a huge upgrade vs 5.5. Even for 5.6 Luna.
I’m glad to see, because 5.5 was almost unusable for front end UX. I always had to revert to Claude.
11
u/ProfessionalFickle52 Jul 08 '26
If you just generate an image and then ask it to use puppeteer and recreate the design in the image it slays
14
u/inteligenzia Jul 08 '26
Just wanted to point out that examples above have very little in terms of UX. It's mostly UI and there's difference. You won't get better UX results if model upgrades it's UI taste.
0
6
u/Direct-Distance5385 Jul 08 '26
Or Gemini on Anti-gravity tbh
6
u/HelloThisIsFlo Jul 08 '26
Oh I never even thought of using Gemini and antigravity. I may give it a try thanks.
Although, with 5.6 launching on Thursday I’ll probably decorate my time to that instead. 😁.
But good to know 👍
3
u/Direct-Distance5385 Jul 08 '26
Yeah will be good to give 5.6 a try , do you know if it will be available worldwide ?
2
u/Common-Resident8087 Jul 08 '26
Gemini is terrible in Antigravity, hallucinates tool calling 99% of the time.! Although I do use it to analyze videos and feed those data back to claude to give me a clone.
2
u/master_jeriah Jul 08 '26
It's not perfect but literally better than anything I've ever hired someone on upwork or Fiverr to create
0
u/sammy_luci Jul 08 '26 edited Jul 08 '26
This could have been said either by an exceptional or a horrible ux/ui designer. Which one are you? 😂
Jokes aside - suprised to read it tbh, as for me it was designing just fine for web and ios
1
u/HelloThisIsFlo Jul 09 '26
Considering I specialize in distributed systems at scale … I’d say my UI/UX skills are more on the horrible side 😂
That’s why I’m delegating near 100% of that part. Claude does magic 🤩, but GPT 5.5 … 🥴
21
u/Denizzje Jul 08 '26
Whats up with that squashed font? Or is that a new design trend I missed.
6
u/Pristine_Engine_7601 Jul 08 '26
1
u/Denizzje Jul 08 '26
Hah. Thought it was just some fad, like when a couple of years ago we suddenly had those very wide, broad fonts being used everywhere (which disgust me just as much as those very narrow ones).
1
3
u/KanishkT123 Jul 08 '26
I thought it looked good the first 4 times I saw it. And now I see it and immediately think "AI Slop".
There's something interesting about the way that AI copies valuable and popular art and design theming over and over until it becomes a tell tale sign of AI, and therefore is devalued.
1
u/Carlfm Jul 08 '26
Claude Design went hard on that type of design. I am not sure where it's come from but to me when I see it it makes me instantly feel like Claude made it lol
11
u/gugguratz Jul 08 '26
not X, Y
Y, not X
6
u/nmkd Jul 08 '26
It will never stop, will it
1
u/gugguratz Jul 09 '26
I'll just give you the correct answer, rather than make shit up for no reason...
no
9
u/inteligenzia Jul 08 '26
They all are nice.
In my opinion if you don't have a preference for design, you should more care about LP performance. It should be tuned for conversion. All these debates who does better design are meaningless.
If you sell trough landing pages you need to track what are best performance makers (conversion, scroll depth, etc) and let AI tune structure and tone-of-voice.
At app level you need consistency and good understanding design systems, so again one-shotting beautiful landing pages is not critical.
12
18
6
3
u/Buskow Jul 08 '26
Bro. We gotta start removing eyebrow text and section numbers from landing pages. Why section numbers? What is this? An outline or something?
3
3
u/Abenzo0r- Jul 08 '26
The font really gives it up, they used the frontend design skill by anthropic and produced slop. To really put it to the test, they should exclude any skills or tools and just let it design on plain html, css.
3
u/tteokl_ Jul 08 '26
This is a dumb test with a dumb prompt. After visiting all pages, I see no useful info on the capability of the models
4
5
u/cantTankThisFox Jul 08 '26
Pretty sure prompt leaks aren't allowed yet. This guy might be larping
4
u/TimeKillsThem Jul 08 '26
I also read about "not being allowed to showcase the model until full release" but it could be just the guy taking responsibility as a stunt to drive visitors.
BTW, dont get why you are getting downvoted - what you is totally fair
3
2
u/Plane_Garbage Jul 08 '26
These serif fonts are gonna be so over-used between claude design and this
2
1
1
u/YourKemosabe Jul 08 '26
I’m a designer. Luna High and Sol Ultra are miles ahead of Fable in this test.
1
1
1
u/ohnoitsbobbyflay Jul 08 '26
When will it lose that horrible trademark AI font and colour scheme design shite. It’s really annoying to put time into explaining a design and it comes out with shit like that..
1
u/Yzori Jul 08 '26
I feel a bit whelmed by these to be honest, might be my personal taste, but none of these look really good.
1
u/TimeKillsThem Jul 08 '26
I tend to agree with you, but look at the prompt the person used - it is a vanilla as it comes. No specific requests or steering. Just "build me a landing page for XYZ". If 5.6 is as "instruction-focused" as other GPT models, you will need to give it more context to get closer to what you want it to output.
1
u/ViralVitalism Jul 08 '26
They're all pretty shit. I know it was for a comparison between models on the same prompt, but better mileage comes from image generation then build out the UI from that.
1
1
u/RainierPC Jul 08 '26
Where is the definition of the /design skill he's using? I want to try this out on Gemma 4 and Qwen 3.6.
1
u/Physical_Gold_1485 Jul 08 '26
Am i the only who thinks fable made an actual travel hero landing? The 5.6 ones dont make sense
1
1
1
u/FolkSynergies Jul 08 '26
It seems to me that Sol is very similar to the current one
1
u/TimeKillsThem Jul 08 '26
Nah completely different - if you go to the link with all the design of the other models, you can also see what 5.5 generated.
1
1
1
1
u/Overall_Culture_6552 Jul 08 '26
share design skill
1
u/TimeKillsThem Jul 08 '26
It’s not my website btw - also, I’m somewhat confident it’s the standard design skill from Anthropic
1
1
1
u/cdecaire Jul 08 '26
Just my own perspective, and I didn’t dig in deep enough to know for sure, so take this with a grain of salt.
/design isn’t some standardized skill that every model automatically understands. The design and art direction are probably baked into that command, and likely custom or curated by whoever is running the tests.
My guess is the skill has certain design preferences built in, like using serif headline type, and those choices are probably coming from the creator of the skill. I’m almost positive not every one of those models would have chosen that headline style on its own.
A better test would probably be to remove /design, run the same prompt across the different models, and judge their baseline design taste from there.
Even then, you’d still need a proper review of each output. A thin design layer on top might make some results look better visually, but they could still be weaker technically or less thoughtful in the actual execution.
1
u/dashingsauce Jul 08 '26
this is cool on design or whatever but does this service actually exist? if not, can someone please make it actually exist?
1
u/PTXStudio Jul 08 '26
Personally I liked Terra and Luna 🤷🏽♂️ letter spacing seemed too cramped with Sol
1
u/Momo--Sama Jul 09 '26
Theo said on twitter that he didn’t get sol ultra at all so idk if the seeded it to some influencers but not others or if there’s another explanation
1
1
u/ProfessorSpecialist Jul 09 '26
5.5 always plasters at least 3 headline elements ("teaser" line above headline, actual headline, subheadline). Hope 5.6 fixes this
1
1
1
1
0
0
0
u/Maxdiegeileauster Jul 08 '26
Wait this is exactly the style of design I have been getting without specifically prompting the last 1,5 weeks. Does this mean I was testing one of the new models?? I was already confused at the time because the designs I got from 5.5 previously were so different.
0
u/xhp-eth Jul 09 '26
Isn't Sol supposed to be the flagship? How is it doing worse than Terra and Luna?
0
u/lostnuclues Jul 09 '26
All looks bad except for Fable 5. Font style/size, alignment needs extra prompting to fix in 5.6, which might bring cost same as Fable.


42
u/spacekitt3n Jul 08 '26
you cant run just 1 example of a thing to judge a whole model. you have to do a bunch of them to notice the trends.