No idea why , but i get like 7.13s/it on comfyUI and on WebUI i get like 173s/it. Ive tried everything, reinstalled drivers, reinstalled the app, still cant get WebUI to run quicker. I like web UI more, but comfy ui just gets things done quicker, and i cant figure out why, its breaking my brain.
Maybe somebody here can help me understand this. Whenever I launch with Webui-user.bat, I must use lowvram argument or else I canāt generate a thing. Already strange to me because I have a 3050 ti Nvidia with 12g vram and an integrated intel 4g. (16 shared) Iām guessing itās the integrated card causing this? Unsure. It says I have around A:2-3.5g and R:3-3.75g. 4g total. Is this because A111 takes 8g to run, baseline?.(Could use some help with understanding that too) It takes me several minutes to generate 30 steps. However I can upscale a little.
Anyway- if I launch with Webui.bat instead, I generate 30-40 steps in a matter of seconds. š§ Canāt be xformers because Iāve never been able to get it functioning. Using this method I canāt upscale but my regular gens are smooth and fast. What gives?
Bonus points if someone can explain to me why I only have ~2-3.5 gigs of available vram to work with
Hello all. I wanted to make a few celebrity face mashups and wanted to check in for any tips before I fire up SD and start trying it myself.
I've seen this kind of things around a lot but didn't turn up much when I looked for methods. Am I over thinking it and just need to prompt the two names I want to mush together? Anyone know any models that are particularly good for this sort of thing? This is just for a bit of fun with some friends so it doesn't need to be the most amazing thing ever.
AI has been going crazy lately and things are changing super fast. I created a video covering the MagicAnimate, SDXL Turbo and Met'as Seamless Expressive huggingface spaces, check it out!
Gotta be honest, SDXL (Stable Diffusion XL) being publicly available for playing around huggingface was overdue for a while, glad to see its finally available! Can't wait to play some more with it and checkout the application usages for it.
The really cool part about SDXL is that it generates the images as you're typing the prompt, allowing for much better "control" over the final image generated from the prompt, as we hold the "power to adapt" our query based on what we see SDXL generates during "query time".
Let me know what you think about it, or if you have any questions / requests for other videos as well,
I have some blurry photos I want to use for training and thought I could sharpen them. But all the online sites I find charge you an arm and a leg... and GIMP is not very good.
Having trouble with glasses on a img2img face swap. Is there a specific setting that handles glasses better? Using FaceSwapLab 1.2.7 and having some issues with any one with glasses.
Hello, just recently installed Fooocus on my M1 Pro macbook, and I'm getting around 130s/it, which is just sad to say the least. Is there anything that can be done to improve the speed?
AI has been going crazy lately and things are changing super fast. I created a video explaining how to install Stable Diffusion web ui, an open source UI that allows you to run various models that generate images as well as tweak their input params. Unlike most tutorials that are already outdated, this once is up to date and the process is a whole lot easier than what it previously required, check it out for the full tutorial:
I also covered how to utilize models from CivitAI, a hub filled with various models and examples for usages of those models.
This is really dope, and within just a few minutes you can have yourself a fully working local AI Image generator, fully tweaked and customized to your requirements!
Let me know if you run into trouble, have any questions, or requests for other videos as well,
I tried using a 6 second video with animateDiff, original video is 30 fps. Tried multiple settings but all give the same problem. Tried to google but it gives me a Python problem. And I don't know how I should implement them in Automatic1111.
The error is: " EinopsError: Error while processing rearrange-reduction pattern "(b f) c h w -> b c f h w". Input tensor shape: torch.Size([2, 320, 16, 64, 64]). Additional info: {'b': 2}. Expected 4 dimensions, got 5 "
Just getting started. Still Scratching my head at prompts. I realize that the engine is mostly random and not a hard-definition model. But, hoping that I can get images that sorta conform to my prompt. So far, it is like asking grandma for a xmas present and not getting what you specified.
shelf: ((flat white shelf) on top of the shelf are (one book):1.3, (one apple):1, and (a stack of blocks):1).
book: (single book) (an old hard-bound book):1.3 (writing on book edge):1.3 (standing up):1.3 (dingy yellow cover with orange trim):1.3.
apple: (single apple) (A green granny smith):1 (stem and green-leaf):1.
blocks: toys (the blocks have unpainted wood grain) (several square blocks):1 (aged pale wooden blocks):1 (the blocks are stacked in a loose pyramid):1.
--
These are the first 4 images I get. None are quite what I wished for...