Skip to content

Comment on Ask HN: Why can't image generation models spell?

Comments

My experience is that whilst it's not perfect, modern models can create images with text that is correct, relatively often.

As a test I just tried ChatGPT with the prompt :-

Hi ChatGPT can you give me a picture of a banner that says "Hacker news"

And the resultant image does indeed have that text on it. Where I've seen this approach fall down, is where the text is long and/or complex or the words are uncommon.

so while there's some way to go, things are definitely improving here.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.