I've been following r/StableDiffusion on reddit for a while and was wondering whether this can also be used for anything that doesn't look like a cheap fantasy or science fiction novel cover.
This is an honest question, I haven't seen any example of anything else so I got to wonder whether the models they are using are specialized for sci-fi and fantasy "air brush/digital" style? Why?
Of course it can. It's extremely new, and people are already creating fantastic results since it has been released literally just over two weeks ago.
The models, processes, and collective knowledge will just develop more over time to create MUCH better visuals, videos, and temporally coherent animations.
This is like a new form of art canvas, and we're all getting used to the basics of using the "paints" and "brushes" for it. In a few months/years, some of us will master the skill and produce fantastic artpieces.
But for know, the stunning stuff tends to be fantasy, sci-fi, impressionistic. Photos of people that are not a portrait are pretty often anatomically impossible. Getting hands, arms and stuff right seems quite difficult for these networks.
Funny enough, I've seen underwater pictures that to me looked quite believable, but to they expert are ridiculous. Lot's of impossible stuff going on. Human brains are ready to fill in a lot of detail.
Interesting. The results so far look pretty good, though only for fantasy and science fiction "fan art" style. That's why I was wondering whether the models are only trained from such inputs. If I understand you correctly, this is not the case and other styles of art can also be produced. Right?
Another question: Do the people who run the software claim copyright on the results even though these are (mostly) produced by the software? It sounds like that when you write "some us will [...] produce fantastic artpieces." I guess it's also legally the case but wonder whether that's also how people experimenting with it understand it.
As it takes a lot of iterations, curation and knowledge about how to best steer the systems, most users (rightfully) feel some sort of creativity and skill went into the works even if the AI did the pixels, except if you're really lucky and you get something amazing out of a simple prompt.. In the end, it's each to his own I guess. This will just be another tool in the toolbox of a digital creator.
Btw it simply isn't true that the AI generators are "only good for fantasy and sci-fi". I guess you've been seeing a biased selection. They can do pretty much anything. MidJourney is for sure more fine-tuned towards artsy stuff though.
Fantasy novel covers are exceedingly easy to do because of the mountain of examples in the training data. Basically: any kind of art that we have lots of examples of are very easy to make with these tools.
My fun has been with two games: 1) making unusual art from prose using the art styles of famous painters, 2) playing "AI Pictionary" with friends (can you produce image X; example: a person eating ramen with chopsticks that are light sabers).
Fantasy novel covers are exceedingly easy to do because of the mountain of examples in the training data
It's also because they are basically nonsense (fantasy) so the results in the style look more plausible.
I tried using a few AI generators to get some basic placeholder images for products and it's utterly shit, it's clear that networks dont understand what they are generating. Like I tried to generate bycicles and I would get components sticking to ground, floating components, stupid proportions, visual artifacts.
Comments
I've been following r/StableDiffusion on reddit for a while and was wondering whether this can also be used for anything that doesn't look like a cheap fantasy or science fiction novel cover.
This is an honest question, I haven't seen any example of anything else so I got to wonder whether the models they are using are specialized for sci-fi and fantasy "air brush/digital" style? Why?
Of course it can. It's extremely new, and people are already creating fantastic results since it has been released literally just over two weeks ago.
The models, processes, and collective knowledge will just develop more over time to create MUCH better visuals, videos, and temporally coherent animations.
This is like a new form of art canvas, and we're all getting used to the basics of using the "paints" and "brushes" for it. In a few months/years, some of us will master the skill and produce fantastic artpieces.
Sure, it will only get better not worse.
But for know, the stunning stuff tends to be fantasy, sci-fi, impressionistic. Photos of people that are not a portrait are pretty often anatomically impossible. Getting hands, arms and stuff right seems quite difficult for these networks.
Funny enough, I've seen underwater pictures that to me looked quite believable, but to they expert are ridiculous. Lot's of impossible stuff going on. Human brains are ready to fill in a lot of detail.
Interesting. The results so far look pretty good, though only for fantasy and science fiction "fan art" style. That's why I was wondering whether the models are only trained from such inputs. If I understand you correctly, this is not the case and other styles of art can also be produced. Right?
Another question: Do the people who run the software claim copyright on the results even though these are (mostly) produced by the software? It sounds like that when you write "some us will [...] produce fantastic artpieces." I guess it's also legally the case but wonder whether that's also how people experimenting with it understand it.
As it takes a lot of iterations, curation and knowledge about how to best steer the systems, most users (rightfully) feel some sort of creativity and skill went into the works even if the AI did the pixels, except if you're really lucky and you get something amazing out of a simple prompt.. In the end, it's each to his own I guess. This will just be another tool in the toolbox of a digital creator.
Btw it simply isn't true that the AI generators are "only good for fantasy and sci-fi". I guess you've been seeing a biased selection. They can do pretty much anything. MidJourney is for sure more fine-tuned towards artsy stuff though.
Fantasy novel covers are exceedingly easy to do because of the mountain of examples in the training data. Basically: any kind of art that we have lots of examples of are very easy to make with these tools.
My fun has been with two games: 1) making unusual art from prose using the art styles of famous painters, 2) playing "AI Pictionary" with friends (can you produce image X; example: a person eating ramen with chopsticks that are light sabers).
It's also because they are basically nonsense (fantasy) so the results in the style look more plausible.
I tried using a few AI generators to get some basic placeholder images for products and it's utterly shit, it's clear that networks dont understand what they are generating. Like I tried to generate bycicles and I would get components sticking to ground, floating components, stupid proportions, visual artifacts.