With some finagling, I got an acceptable image, but it was much worse than I would have expected. On the other hand, the whole thing made for a funny story to tell when I was handing over the present.
This is line art (because I wanted to print it in black and white), but the non-line-art style had the same issue. Also it's Saint Nicholas and not Santa, but I just confirmed that the issue still persists with Santa.
And A woman giving santa a massage is just too similar to a bunch of porn...
Further fine-tuning would certainly improve the final result.
ETA: They still can't seem to make it with the white tie, though!
Open it up in GIMP, paint the shirt blue and jacket red.
Feed resulting image into img2img with a low denoising strength (0.3-0.4 maybe) with the textual prompt of what you want.
This can help you get combinations of things the AI doesn't easily do itself.
I can see many types of clothing could completely throw off the cars in a crosswalk type situation. Reflections are solved with lidar, I wonder if a sudden breeze picks up a dense cloud of dust, that the car reacts very unpredictably.
Its a possibility that insurance scams will evolve in the future to target these vehicles using adversarial techniques that for humans are obvious. How could we even stop that? It'd likely be a cat and mouse game.
I think the reason is it doesn't understand compositionality, meaning it doesnt have any logic or "intuition" about the relationships between the colors and items (or eyes and forehead) so it fails to constrain the way they are assembled in the image.
No Shiva images were included in the training data?
https://stable-diffusion-art.com/how-stable-diffusion-work/
It’s a quick read and I found it very helpful.
Remaking old computer graphics with AI image generation - https://news.ycombinator.com/item?id=34212564 - Jan 2023 (73 comments)
Six fingers, yeah, and also car wheels are a random mess of spokes and bolts.
Best horror film of last year.
https://www.reddit.com/r/deepdream/comments/v9jr73/how_can_y...
This is really funny actually, considering what basic Photoshop tools are capable of out of the box :)
1. Generate me an eye
2. now place 3 of these on a head where I choose
3. merge it in for me, in some way that fits the overall image.
The 4th step (missing due to time it takes) iterate over a few iterations till I see one I like.
3 years from now a novice will be able to do more advanced photoediting than a pro inside photoshop right now.
It's already the case that photoediting for novices are a reality, but the level of control is lacking.
https://lh3.googleusercontent.com/OFIpc-iFtO1AQty-IJ6beLEGe6...
He already had a good enough lantern, all he had to do is to set a tasteful setting for this.
https://imgur.com/a/SgQ38bw https://imgur.com/a/Gi0IZtB
Can you apply feather edges for me and compare the results (and let's use the gentlemen's agreement that you will not spend more than 20 seconds)?
Model output being defined by the training data is still a valuable observation for many unaware of the limitations, regardless of how obvious it might be for others.
But half the fun is trying to do it entirely inside the generative art tools available. If I were really trying to make something good I'd use both together.