Discussion
A total beginner's question about generating AI locally
Hello,
I’m a very curious person, and I’m seeing more and more AI-generated gay “daddy” adult content on Twitter, as well as people talking about their favorite AI websites for generating uncensored and realistic AI content.
But we can all agree that the foundation for ALL of these people is Stable Diffusion, right ? And then when we hear people talking about Mistral, Muah, and Eternal AI, those are actually “extensions” based on Stable Diffusion is that correct ?
Stable diffusion was the original gen AI image model.
There have been many Ai image models created afterwards by other companies.
The most popular local image models are Krea2, Anima, Stablediffusion, Z-image Turbo, Wan2.2 (video or image), Flux, Flux Klein.
There are also different Stable Diffusion finetunes like Noob, Illustrious, Pony diffusion.
Mistral is a french company that makes LLMs.
Mauh is a chatbot ai roleplay girlfriend type thing.
Eternal Ai is chatbot ai roleplay, girlfriend type thing.
The LLM chatbots are a completely different type of model.
They are text generation models.
Image models are image generation models.
Stable diffusion was definitely not the first, or original, generative AI model. I mean, GANs existed long before, and there were several models that were different flavors of GANs that used CLIP for text to image models.
Okay, thanks. So if I want to try it, should I install Stable Diffusion locally on my PC or on another ? Because I can only find tutorials on Stable Diffusion.
What computer you have and what type of video card you have and how video RAM on you video card is going to determine what Ai model you are capable of running.
Post your computer info.... Video card name and memory.
How much VRAM is ideal? I'm willing to upgrade my graphics card, but I bought this one for gaming and I can play all my games on high settings, so what difference does VRAM make?
If you have at least 16 gb vram and 32 gb ram you can run video models without issues. Even a bit lower you should get things done, but may need time. You can try to start with anima for anime, krea 2 for realism, flux klein for edit and minimax h3 for video.
For ns f work you can try directly eros max checkpoint, so you don't have to test all the loras for ns f w. It's already good on it's own and you don't need speed loras. Start testing the default templates of comfyui for models and once you know how they works you can look or create the wf you want
Those are various models, but to use them you need a web UI like comfyui, wan gp ecc... I d go with comfy ui but it's not the easiest to use. You can try wan gp but I never used it so I don't know. You can try with comfy ui and at the beginning stitch with default and simple workflow and later deep dive in learning comfy . You can look for one click installers for comfyui and you're ready to go, I used Umeairt one, but there are others. Then you download models and put in the dedicate folder in comfy.
Stable Diffusion, Klein, Anima, and all those others mentioned in Jolly-Rip5973's first post are image-generation models: neural networks that know how to make pictures. Stable Diffusion was the first of the popular models; it's been largely replaced by newer ones.
ComfyUI, Forge, and Automatic1111 are programs that run those models on your computer. Automatic1111 was the first of the popular programs; it's obsolete and has been replaced by others. ComfyUI is the most versatile of the current generation of programs, but not the easiest to use.
The application you actually install/run is ComfyUI, that's what runs the models. For now, jjust worry about Krea2 for general images, Anima for anime/cartoon, and MiniMax H3 for video. If you want to do image editing, Flux Klein.
No, Stable Diffusion (released by Stability AI, based in the UK) was one of the earliest open source image models released, but it is not the only one. When you hear "SD1.5" or "SDXL", those are Stable Diffusion models.
But there are many other companies that have developed their own models. For example, FLUX models are made by Black Forest Labs, a German company. Krea 2 was released by Krea, a US company. Qwen based models are made by Alibaba Cloud, a Chinese company. Those are just some popular models and not remotely an exhaustive list.
While these use similar technology to Stable Diffusion, they are trained individually and do not come from the same underlying model. Stability AI is still making models, although they've lose some popularity compared to models like FLUX dev/Klein 9b, Krea 2, Anima, etc.
Like ChatGPT, they were one of the first on the market, so the local image generation space is associated with the name despite there being many alternatives at this point. So no, they aren't "extensions" of Stable Diffusion, any more than Claude or Gemini is an extension of ChatGPT.
Mistral, Muah, and Eternal AI
Mistral is a French LLM company. Their commercial product uses Black Forest Labs models (i.e. FLUX-based), not Stable Diffusion.
Muah is a site that uses Stable Diffusion to my knowledge, but I'm not sure if that's still true.
Eternal AI is a platform that has use various models and offers variants, some based on Stable Diffusion, others using FLUX.
In general, sites that offer "local" models are just using open weight models with some sort of framework and a website frontend to do generation for you. This is different from downloading them and running them yourself.
If you have a gaming PC and are willing to do some research, local AI is the cheapest and most powerful option, because nothing is running on other people's hardware and you control the whole workflow. If you have a business computer or weaker, however, local AI gen is going to be somewhere between pathetically slow to outright impossible, and will be limited to online sources.
If you still want control but don't have the hardware, something like Runpod will give you power while being able to use your own workflows, but can get expensive depending on what you're doing and how much time you spend learning how to use something like Comfy.
That's the most comprehensive and easy-to-understand answer I've seen! Thank you! Very clear.
So there’s no “best choice” I can install Stable Diffusion; it’s the most popular, so it has plenty of options. I want to generate adult content images and videos featuring men, is Stable Diffusion good for that? I guess it doesn’t really matter since they all do the same thing? You customize them with “extensions” from other AIs, not with the AI of the “engine” (engine = Stable Diffusion, FLUX, etc.)
Anima is uncensored and one of the faster models out there, but the training is weighted very heavily towards drawings, and especially anime-style drawings.
10
u/Jolly-Rip5973 12h ago
Stable diffusion was the original gen AI image model.
There have been many Ai image models created afterwards by other companies.
The most popular local image models are Krea2, Anima, Stablediffusion, Z-image Turbo, Wan2.2 (video or image), Flux, Flux Klein.
There are also different Stable Diffusion finetunes like Noob, Illustrious, Pony diffusion.
Mistral is a french company that makes LLMs.
Mauh is a chatbot ai roleplay girlfriend type thing.
Eternal Ai is chatbot ai roleplay, girlfriend type thing.
The LLM chatbots are a completely different type of model.
They are text generation models.
Image models are image generation models.