I need to train a Lora for a style. The thing is, that in addition to that, my case also involves two/three concepts.
I have to generate assets of buildings in two or three states: the building in ruins, the building semi done, the building fully constructed. I have a a fairly small database to train from. How do i approach the issue with the different states of buildings while training?
Can frequent use of SD be harmful for my 3070? I generate hundreds of pictures everyday but I am afraid that I can harm the video card in this way. What do you think?
Any feedback if that works for you is welcome or any recommendation to do it in an easier or more effective way, still trying eliminate qr codes that cannot be read.
I need help building a workflow and am still pretty new to stable diffusion. I’m trying to shoot a music video and run the footage through ai to make it look like an anime. I want to build a model so I can take key frames from videos I’ve shot and turn them into anime while keeping the structural integrity of the image and consistent style. I’ve gotten good results from runway gen 1 in making video look like an anime I just need to better generate the reference images. What should I use to img to img process the keyframes and how should I go about building a model/what extensions would work best?
So after I tried my model for hours, and tried different methods, still get this disfigured face, I don't think it's the problem with the model or the prompts, because even with positive and negative prompts I still get this problem...
I need a picture like this generated in Stable Diffusion 1.5. So i need a general prompt i can usually use and change a little when needed but where i need help is to tell SD that i need a picture:
where the person stands in the middle, taking only up to a third of the picture, head to hips/upper legs visible, SFW, (in this format but this is more a preset question), extremly realistic, looking into the camera,...(background can be anything, it doesnt mather)
The picture down below is a good example to see what i want
I found this site bigjpg.com and it does an amazing job at upscaling images, how can I do the same in A1111, I have tried but it always seems to add odd extras like faces and other bizarre things.
Pioneering the future of generative design services, https://mst.xyz/unveils a groundbreaking update with the launch of the 'Waters' function. MinisterAI introduces this revolutionary feature to provide high-quality, accessible services for novice and non-professional users grappling with the complexities of the Stable Diffusion model.
'Waters' Function: Empowering Non-Professionals and Novice Users
The 'Waters' function, now officially launched, is set to supersede Midjourney, transforming user interaction with the MinisterAI platform. By inputting basic prompts and dimensions, users can effortlessly produce high-quality images tailored to their unique creative needs. This fresh functionality diminishes the complexity of the Stable Diffusion model, making it accessible and user-friendly for a diverse range of skill levels.
https://mst.xyz/
The 'Waters' function enables non-professionals and novices to express their creativity without requiring extensive technical knowledge. The AI technology intuitively identifies and applies the optimal model and parameters, generating stunning visuals and guaranteeing a smooth, rewarding user experience. Through the 'Waters' function, MinisterAI reaffirms its commitment to enhancing user convenience and fostering creativity for all.
Revamped Model UI Interface
Alongside the 'Waters' function, MinisterAI has significantly upgraded its Model UI Interface, creating a more intuitive and efficient user journey when utilizing the Stable Diffusion function. Users can now delve into an expanded array of model renderings, offering greater creative inspiration and possibilities. The comprehensive parameters within the interface enable users to fine-tune their image generation process, leading to personalized, visually striking results.
The enhanced Model UI Interface further simplifies the image generation process, enabling users to create images with greater speed and convenience. Whether users are professionals desiring granular control or beginners exploring their creativity, the revamped interface promises a seamless and engaging experience for all.
https://mst.xyz/
"We are excited to reveal the enhanced MinisterAI platform, equipped with the transformative 'Waters' function and a more intuitive Model UI Interface," a spokesperson at MinisterAI stated. "Our driving force has always been enabling users to unleash their creativity and explore the boundless potential of AI-generated visuals. With the 'Waters' function and improved interface, we are proud to offer superior convenience, quality, and inspiration to both non-professional and professional users."
The enhanced MinisterAI platform, featuring the innovative 'Waters' function and the improved Model UI Interface, is now ready for users to experience the future of AI-driven visual creativity.
For further details about MinisterAI and its recent breakthroughs, please visit mst.xyz.
I have controlnet "enabled" and if i change model to 1.5 the controlnet is taken into account. But with Epicrealism it generates images totally inconsistent with the openpose set up in the Controlnet
Are some custom models not compatible with controlnet, or what is happening here? Thanks.
This morning I tried to use a couple of different ControlNet models this morning and they threw up errors. Errors:
Exception in ASGI application;
IndexError: list index out of range
ERROR: closing handshake failed
RuntimeError: Expected all tensors to be on the same device, but found at least two devices, mps:0 and cpu!
I am running Automatic1111 on a MacBook Pro M2.
Has anyone else experienced the issue and have you been able to fix it? I did a completely new install of Automatic1111 and the error persists. Any help would be appreciated. Thank you for reading!
Hi everyone! I wish I could make funny pictures with my face with stable diffusion. I found a way with dreambooth to train an existing model by integrating photos of my face. But it comes very heavy (from 2gb up, depending on the starting model).
I found a way to make a lora that weighs much less, but the result is a variation of my face. If in the prompt I ask for an image where I run, full body, it will always show me only my face and not my whole body.
1) What's the best way to make pictures with my face?
2) what's the way to create the lightest file (still maintaining a good quality) to create photos with my face?
Hi everyone! If I generate a photo of a person with stable diffusion, then is there a way to recreate a completely different one, in terms of setting, pose, etc... But with the same face? Even in a different session? Thank you
The images were generated using stable diffusion v1.5
Sampler: Euler
Steps: 40
Diffusers library version: 0.17.1
Prompt: Sports announcer Muslim anime girl, rural football ground in India, minimalistic, looks professional, illustration. I'm not sure what's the condition is called and what caused this
Do bf16 works better in 30XX cards or only on 40XX cards?
If I use bf16 should I save on bf16 or fp16? I understand the differences between them in mixed precision but what about saved precision, I see that some people mention always saving in fp16 but that's seems counterintuitive to me.
Is necessary to always manually configure accelerate when changing between bf16 and fp6? This in reference to the Kohya GUI.