Skip to content

Comment on Ask HN: Why can't image generation models spell?parent

Comments

Faces don't have lots of repeating similar subcomponents beyond some things that are just two items in bilateral symmetry (teeth are a big exception, and teeth, when visible, can be a problem.)

And, actually, faces still, especially outside of closeups of just the face, can be a problem, too, which is why a separate face restoration with a GAN or inpainting pass for faces with the same or different diffusion model is common.

faces are simpler since unlike hands most of their major constituents are at fixed relative positions to each other. but the flip side is that people are hyper-biased towards attending to facial details, hence why they were basically the first-handled special case

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.