Helpful Resources
I’ll add more here as I remember them. Feel free to add more in the comments.
- AUTOMATIC1111’s Stable Diffusion WebUI is the software nearly everybody uses
- System requirements wise, for a while I used it on a 1050Ti with 4GBs of VRAM. I wouldn’t recommend going any lower than that. An RX580 with 8GB VRAM does wonders at a similar secondhand price point (if there isn’t any crypto hype going around where you are)
- Using https://github.com/pkuliyi2015/multidiffusion-upscaler-for-automatic1111 can provide a really nice speed boost if configured correctly.
- Civitai has a really good selection of models, loras, and other resources
Models
Models are basically the brains of Stable Diffusion. They are the data SD uses to learn what your prompts mean.
The built-in models that come with Stable Diffusion are really bad for porn. Don’t use them. In fact don’t use them at all unless you’re training your own models, there are better SFW models.
Here are some of my personal favourites:
Anime
- MeinaHentai is a great model to start with. Compared to other models it’s really easy to prompt
- AOM3 also does really well, though it might be a little more difficult to guide
For all of those, I recommend installing https://github.com/DominikDoom/a1111-sd-webui-tagcomplete, as they heavily rely on danbooru tags.
- Berry Mix (Pre-mixed version here) can also work pretty well, depending on what you want to do. AFAIK it uses rule34 tags instead of danbooru, so it probably won’t work all too well with prompts used for the above ones
Realistic
- Uber Realistic Porn Merge is the only realistic model I know of that does hardcore stuff. It’s unfortunate problem is that it’s REALLY DAMN HARD TO USE
VAEs
VAEs are mostly used for finetuning colors, sharpness, what have you. Some models come with a VAE builtin, but for ones that don’t, it’s recommended to have one on hand.
- “Anything VAE”, “Orangemix VAE”, and “NAI Leak VAE” are the same exact thing under different names. If you already have one on hand, don’t bother with the others. Most VAEs are renamed versions or modifications of this one.
- Waifu Diffusion’s kl-f8-anime2 is also a pretty good one. It doesn’t require Waifu Diffusion.
- The one that comes with Stable Diffusion is the only one that seems to work for realistic stuff.
LoRAs
LoRAs teach models about concepts (characters, clothing, environments, style, …) they might not know about. There are a LOT of them, so feel free to browse Civitai to find ones you might want.
LoRAs tend to be specific for families of models, or at the very least styles (using anime LoRAs on realistic models tend to be a bad idea), but there are a fair few that will work across the board.
Locon and LyCORIS are newer formats of LoRAs. Not sure on the technical differences between them, but they will not work out of the box and need an extension such as https://github.com/KohakuBlueleaf/a1111-sd-webui-lycoris to get working
Textual Inversions / Embeddings and Hypernetworks
These are mostly obsoleted by LoRAs. There are a few embeddings such as Deep Negative and EasyNegative that are still quite useful, but in most cases you’ll want to use LoRAs instead.
Is it just me or do all the AI generated porn images have something in common that makes them immediately recognizable? I am not quite sure what exactly it is but they do all have some quality in common.
Its the uncanny valley effect I think. Your brain recognizes that something is wrong in the image.
That might be part of it but I think there is also a sort of sameness, certain aspects that are always the same in the generated images and more varied in real ones. Possibly related to the averaging of all the input data.
Did anyone manage to find out what model sites like anydream.xyz might be using? I have hard time emulating the style…
Interesting model if you are into titfuck: https://civitai.com/models/7701/realistic-titfuck?modelVersionId=9042
Is everyone creating accounts with CivitAI to get these models?
I did, it’s definitely worth it
Why did mine get downvoted to oblivion
So what settings do y’all use when playing around with prompts, before you generate a few really nice HQ ones?
I use 600x800 (or 800x600) and 50 steps when experimenting with prompts. Then when i get one i like, I lock the seed, maybe do a few .01-.03 variations. Then lock the variation seed and turn on 2x HighRes Fix. That outputs a 1200x1600 that very closely looks like what I expected. (I’m using a 3090 with 24GB vram)
In my experiments, i found that doing quick and smaller tests with a low number of steps, then increasing it for the highres would change the output too much. I settled on this so that theres less options to toggle back and forth.
What’s your setup?
Are these softwares all free to download/use? Also, how does one start doing this? do i just need the WebUi or do i need extra files to feed into it and stuff?
I just figured it out tonight playing around with the links and readmes available above. If you get stuck I can try to answer more specific questions.
Hmm ill probably have mroe questions but for now im curious:
- how much space was the download(s)?
- how confusing is the software to use?
- what kind of limitations does the software have? can i do multiple people? monsters? futa? etc.
Thanks for your help! :)
Sure,
1: The initial download was pretty small ~10GB. But with the models, lora(s), extensions, ect. I’m up to ~60GB.
2: The guide at the top was pretty easy to follow. Install the dependencies, then install the UI. Launch and run. There is a bit of a learning curve with all of the options but so far it hasn’t been too confusing.
3: That’s where the extra models/lora(s) come in. Various models are trained in different styles, poses, actions, ect. The lora files are smaller things, like poses. IE: Cowgirl is it’s own lora file that tells the model how to use the prompts you give.
Yea the webui and stable diffusion models/checkpoints are generally free to use.
I recommend following a setup tutorial on YouTube. They usually link to models and stuff youll need. Look for a recent one, this stuff evolves a lot.
I’ve been meaning to figure this out. So the commonly used ones are labeled WebUI, but just how much of the content goes to the web itself?
If I wanted to train on images that I don’t want going online, will they? Or will the products that I create end up online, or does all this stay local?
By Web UI it means that the graphical part of it – where you write your prompt and hit generate – is running inside your browser and not as a separate window or command line. Everything is kept on your own computer unless you explicitly tell it to open up remote access.
It was probably a dumb question, but I wanted to clarify just the same, thanks. When ya get these kinds of pics under strict confidence, you don’t wanna risk breaking that in any way.
It was probably a dumb question
No such thing as a dumb question. but yes the WebUI only exists so you dont have to type into a command prompt. It looks nice and pretty for humans, and then translates all those settings into a command that is then sent to stable diffusion. All of it stays 100% local, unless you go into the WebUI settings and tinker with remote access. But it is off by default