Skip to content

Comment on Ask HN: Why can't image generation models spell?

Comments

What is spelling but putting into letters what is heard? If you can give it the text, then it's not spelling is it, it would be copying! I think you want to give it ideas, that will translate into words, which is certainly not spelling. It's creating. SD models start with simply static, and it's asked to find some pattern and expand upon it until it matches the pattern better. Letters on a sign for example are not right or wrong by their placement, but upon judgement by someone who has the knowledge of the language. Sale might mean 'salt' if you speak italian, but it might mean there's a discount for a short period of time if you speak english. Sensibel might seem like a misspelling if sensible if you speak english, but it's perfectly correct spelling in German.

My suggestion is to use Image to Image, start with the text of your son's name, and give it some gaussian noise background, and then paint out the parts you want to keep.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.