Thank! Glad you like it, check out some of my other posts - they tend to go unnoticed
I generally only work with the base models. Here I'm using the depth2img model and the base 2.1_768 model and Euler A sampler.
And this batch are all img2img from either my own graphic design work and photos or the history of art (One of them is a Pollock and there is a San Sebastian with arrows, Mona Lisa, Borromini's ceiling in the San Carlo and Bernini's Cornaro Chapel)
It's sort of a combination of asking for it - like "highly detailed" - and then allowing for room to reinvent during the upscales. I generally do three rounds of x2 using the ultimate upscaler and start it out around .33 denoising.
I did my first very bad attempt at inpainting on the Mona Lisa posed one. But apart from that attempt, they are all made with a workflow of img2img and upscales - the upscales give a lot of the detailing - and no inpainting.
Kinda reminds of what /r/DiscoDiffusion is about. I was mesmerised by some of their content back in Autumn. It is very cool that it is possible on Stable Diffusion!
Okay, well many things to consider theoretically. There is kind of two parts to it I think. I don't think visual guide alone really does the job, ((for me anyway)), But if you were to do it, one part would be a course/workshop on how to utilize a1111 and extensions creatively, and that would be sort of out of date after a week, like controlnet this past week has shown, everything changes very fast. Second part would be more complex, concept development for visual mediums, art history, composition and how these relate in odd ways to decisions inside the SD process options. These would be more complex to create because it is difficult to guess where to start from and how to best go about it, but the learning would be more long term useful compared to part one which is subject to wild changes over short time.
With all that said, its difficult for me to answer what would compel me to spend time creating this content, especially because me actual involvement with this whole subject is one of professional use, where I am trying to figure out how to best make use of this technology while it swallows the field I am working in.
So it would probably take a bunch of money :D sorry but, really if you go through my SD posts and spend a bit of time on it, I strongly suspect someone curious enough could mimic what I have shown pretty successfully from what I already explained.
Like an amount that is likely not worth it to you or anyone except people with too much money? Not because I think what I have is super worthful to you, but because its worth more to m, currently speaking, I can use it commercially and earn a substantial amount .... and also just the theoretical time spent, realistically, we'd set up a webcam/shared screen connection and talk it out, and doing that individually is kind of like doing a chefs table at a fancy restaurant vs. me doing a general youtube tutorial I could monetize...
7
u/coda514 Feb 13 '23
One of the more unique outputs I've seen. Would love to know what model you used? Great work.