ZH version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
67% Positive
Analyzed from 2083 words in the discussion.
Trending Topics
#design#things#aesthetic#https#image#something#though#more#emoji#aesthetics

Discussion (57 Comments)Read Original on HackerNews
Something to consider regarding the narrow space AI-created designs align on: LLMs are trained to write consistent code. This makes sense for something like a billing or a backend function - you want that code to be consistent. The problem though is that LLMs write code to represent designs as well, which means you get consistent designs. You end up aligning on a generic mean because of it, which is often why you see these repeated aesthetics.
It was something I saw when I was working on AI tooling over at Figma. It was very hard to get creative, unique outputs out of LLMs.
One thing I'd recommend if you're trying to avoid this is try a diffusion model as a starting point. Gpt-image-2 is a VERY capable designer, and Opus and Fable are fantastic at converting images to webpages. Starting with images will let you sidestep a lot of the uniformity of output that LLMs have. I'm heavily biased here as this is what I left figma to build (https://news.ycombinator.com/item?id=48995754), but even starting with gpt-image-2 to give you a general sense of the look/feel will heavily differentiate you on the design front.
Here are some examples of image->webpage outputs I've been playing with:
https://html.non.io/tarot/
https://html.non.io/neonRamen/
Stuff is jumping around after I try to move the canvas around, unexpected stuff. I can’t figure it out.
I literally just got my SOC2 compliance complete an hour ago, and I'm gonna roll out team plans/enterprise plans next week (unified billing + shared brands). For teams I'm gonna charge a flat 25% margin on API costs, and for enterprise I'll have a 50% margin + minimum spend.
I think I'm going to keep the individual plans at 0% though. Realistically until my own diffusion models I'm training are the main model I'm using for generation, I'm just an API wrapper + harness. At the individual level I want to be competitive with going directly to the api and coding your own harness.
I am curious though - what would you pay per image?
But basically diffui's build step provides an API to create normal/roughness/depth maps from images. It's essentially an interface to fal.ai's Patina model, which does a really good job at creating those.
The one thing that's missing with it though is metalness maps (Patina can theoretically output them, but they only have a ~10% success rate I've found). I'm trying to train my own diffusion model now to help generate those so the reflectivity isn't fully uniform. I just posted very, VERY early results here - right now my model has around a ~25% succes rate: https://x.com/pwnies/status/2082980850120163756
On the Tarot one, the 3D card in the bottom is both very elaborate and also pointless. No human would've put effort into making that. Not bad per se! But AI.
The reason I feed in the json blob is it lets me better preserve subjects across multiple screens (ie you probably want to carry over the header exactly as it is in one image to another image).
The example sites above though were pretty simple. I think they were something like
"A tarot card website with an inky, painted artistic style, with rich illustrations on a card featured on the left, and a button to draw cards on the right"
and for the other it was something like,
"A cyberpunk inspired ramen cart website"
The funny thing is that looking like everyone else was once considered the definition of good UI design, because it helps users understand a new interface. Then mobile apps came and suddenly UI design was about “expressing your brand” or whatever.
I actually think the other things that Jim mentioned, such as the async task gradient / standard sidebar for agents are good things.
That was really a test of whether or not newer models could implement box shapes that aren't provided by CSS. Opus 5 was one of the first that really nailed that - it's been a long standing test I've had for img->html.
First, they took my em dash. Now, they’re taking my neutral background with orange accents.
(Randomly, I love the Solarized colour schemes, so anything similar is great, vibe-coded or not.)
If the agent is set up to properly show intent then this is better than spinners or exponentially decreasing loading bars.
“Two ways,” Mike said. “Gradually and then suddenly.”
— Ernest Hemingway, The Sun Also Rises, https://standardebooks.org/ebooks/ernest-hemingway/the-sun-a...
Speaking of this, GitHub replaced theirs with a pancake emoji.
Maybe related to the stacked pull-requests announcement earlier today.
[0]: https://github.github.com/gh-stack/
The real question is what the post-AI aesthetic will be, and at a guess it will bifurcate into obsessive recreations of earlier happier pre-AI times, and some sort of opposite extreme of videos of Chinese cities covered in LEDs.
Fable can easily generate something like a design system for itself from whatever starting point you give it. So if you want different aesthetic, you can bang out the aesthetic first in Claude Design and then serve it to Claude Code to follow. It will follow.
For me it was a lot of work with Fable to keep it focused on producing a minimal, medium-high information density system. Sure, it works very well if you are just IP laundering another product, less so for more niche applications. Still super useful, but I haven't seen it produce much "intentional" design on its own.
Seeing Bootsrap for some custom software at my university pre-llm: "I think that a programmer made this and didn't have a designer helping them so they just used a library"
Seeing vibe-coded frontend: "This app may be entirely vibe coded so I am deeply suspicious of it, or it could be a programmer who didn't have a designer helping them and thus used Bootstrap"
It generally stood for either:
- fireworks (combined with other firework emojis,)
- birthdays/festivals (next to the birthday cake and "party popper" thing emoji,)
- magic (next to the wizard or wand, which features almost identical sparkles)
It was pretty festive. Now it's just AI, lol.
The bubble in tech is pretty strange but I digress
On a more serious note, I like the tiny icon trend as depicted in the article. Those little icons next to menu text are pretty ornamental, and don't need to be as big as they usually are.
I'd barely heard some of these words before, and now they're everywhere. It's like everyone suddenly has a friend who speaks in the same idiosyncratic way, and we're all picking up their vocal stims.
In the english speaking world the political sides are even becoming unsubtly linguistically distinct. You can tell which way someone votes by their choice of vocabulary.
Ah, unicorns. We've seen plenty of them throughout the AI bubble. Very apropos.
Midjourney has a definite aesthetic.
Minimax has one. Stable Diffusion definitely (ugly, but free )
Sora everything is rendered in Octane lol
If it’s cartoon or animated it’s Ghibli.
I have not noticed “orange” as a particular AI trend.