I don't know much about the technical aspects of text to image diffusion models, so the fact that I tweaked a local model that's 70% close to what i want already.... that bodes pretty well for what's possible a year or two from now, especially if it's in the hands of someone who
does know what they are doing. For me unfortunately I think the model is a bit overfitted and I didn't caption my dataset smartly enough. It tends to fix to the same camera angle even when i try prompting otherwise. Probably not enough diversity among my training images. See I don't collect spanking images nearly as much as I do videos, so I didn't have a huge training dataset to build from.
If anyone knows where I can download a huge trove of spanking images from let me know . I mean I know there's this forum, but it seems mostly for video uploads.
But enough yapping here's the results.
You must be registered to view images
You must be registered to view images
You must be registered to view images
You must be registered to view images
You must be registered to view images
You must be registered to view images
Still lots of improvements to be made. Will keep working at it.