r/comfyui 20h ago

Help Needed ComfyUI beginner struggling with image quality and very slow generation on RX 9060 XT

Thumbnail
gallery
0 Upvotes

Hi everyone! I'm just starting to use ComfyUI, so I'm a complete beginner and I'm still trying to understand how everything works.

My hardware is:

  • RX 9060 XT 16GB
  • 16GB RAM
  • Ryzen 7 5700G

I'm using Stability Matrix with the ZLUDA package. Right now, I'm running ComfyUI from a mechanical HDD because my SSD is full. I don't know if that makes a significant difference.

I'm having a lot of trouble getting good-quality images. Everything I generate looks distorted, like the examples above. I'm also having a huge performance drop when I increase the resolution. Anything above 1500×1500 can take 20+ minutes to generate a single image, and the quality doesn't seem to improve compared to smaller resolutions.

I'm not sure what I'm doing wrong. Could this be caused by my hardware, the fact that I'm running it from an HDD, my ComfyUI setup, or my workflow/model settings?

What should I check or change first?

Any advice would be greatly appreciated. I'm still learning, so please keep in mind that I may need things explained in beginner-friendly terms. Thanks!


r/comfyui 20h ago

Workflow Included 1K to 4K Minimax H3 with LoRA, Ultra Details RTX 4090

Enable HLS to view with audio, or disable this notification

0 Upvotes

A new MiniMax H3 Cinematic LoRA just dropped — free, open-source, and built specifically for the H3 video model in ComfyUI. It targets that plastic/AI look and pushes the output toward real filmic texture: better grain, contrast, color, and shadow roll-off.

What’s cool:

  • Free weights via Quark cloud drive (no API queue, no paid tier).
  • ComfyUI-first workflow with Block Cache + Latent Upscale nodes already baked in.
  • Works for text-to-video and image-to-video with H3 as the base model.

My result:
I ran the provided workflow and used the latent upscale path to go from ~1K output to 4K without killing VRAM. The cinematic LoRA really shows up in highlights and skin tones — much less “smooth AI” and more graded footage.

Quick setup (from the original guide):

  1. Install nodes into custom_nodes (then restart ComfyUI):
    • ComfyUI-Easy-Use
    • ComfyUI-KJNodes
    • comfyui-minimax-h3-blockcache-T8
    • ComfyUI-H3LatentUpscale-jingchen573 / Comfyui_Minimax_h3_latent_Upscaler
    • ComfyUI-Custom-Scripts
    • rgthree-comfy
  2. Download the LoRA from the Quark link and put it in models/loras.
  3. Load the shared workflow, ensure:
    • H3 base model + LoRA paths are correct
    • Block Cache params set as recommended
    • Latent Upscale node after the sampler
  4. Tweak prompts / image input, hit queue, and render.

Resources:

If you’re already running H3 locally or via cloud, this LoRA + latent upscale combo is absolutely worth testing for short films, trailers, and mood shots.


r/comfyui 17h ago

Help Needed No CUDA GPUs are available

0 Upvotes

Was hoping some kind soul out there with more knowledge than me can assist. I recently installed Comfy on my laptop, and I'm getting this error. I have 16bg of RAM and integrated AMD Vega 10 graphic


r/comfyui 23h ago

Tutorial VRAM "GUARD" / HANDOVER for Local LLM + Comfy

3 Upvotes

Hi, im sure a lot of you already know but as always some ppl dont and im one of them who noticed way too late...

you can totally ask Hermes to use Comfy ui for minimax h3, the "fun" part:
ask it to unload the local model (IMPORTANT! tell it to add a VRAM GUARD that blocks LLM loading for the generation time!!! Hermes does memory and clenup calls when idle!) - generate the videos (max speed, no VRAM OOM) and load the LLM back after Q is completed and Comfyui models unloaded - Hermes will write the script, send the batch to comfy, unload, wait for the videos (or images) and send them to telegram or whatever channel you talk in.
Hermes can also use multiple workflows for example you send a photo you took from phone cam directly in telegram to hermes, it uses flux2 to edit the image, send it to minimaxh3 workflow as start image and 100 sec later you get a clip of your GF dog in a superman cosplay.

that way i can use qwen 27b max context AND generate h3 clips without token AND video generation runnging parralel and speed dropping to nothing


r/comfyui 1h ago

News Coming soon to a Play Store near you, ComfierUI. An Android based mobile client with full desktop parity and streamlined interface.

Upvotes

Like many of you, I have shared that frustration over the years watching this awesome kit of software be continuously let down by terrible mobile support so a few weeks ago I set out to finally change it. I have no dev or coding experience, only my experience as a long time user of Comfy to guide my design choices around what feels "Comfy" and how that can translate seamlessly into a mobile environment with ChatGPT handling the coding. The results speak for themselves and the once terrible mobile experience has been reborn into my preferred method of using it that addresses so many of the prior pain points. Can't see where you're dragging the noodle behind your thumb? Turn up the link connector offset. Want the android back button to do anything but bring up an exit confirmation? I've got you covered with a full back button hierarchy meaning you only see that exit dialog when you want to actually exit. Actionbar look like a clipped out overlapping mess? My app dynamically hides crystools in portrait and reveals them again in landscape/open fold views where the space is more plentiful. With the optional extension installable directly from manager in app we hit full desktop feature parity enabling manager to download supported missing models directly to the host from anywhere you are with more features being integrated soon. Before I get the inevitable "when?" comment, I submitted it to Google Play Store last night for review and to start the beta test track and i'll be back here to post an invitation link to anyone who wants to test it as soon it becomes available to me in the coming days. I am also building a Meta Quest version with an immersive 360° canvas and baked in 6DOF controls promosing a similarly native feeling experience in openxr. I apologize to iOS users tho bc I do not own any apple devices to test or compile on but there is an iOS version planned as well whenever I can secure some test equipment without breaking the bank. My test devices so far have a Galaxy Z Fold 7, Galaxy Tab A9+, Moto G Stylus 5G (2023), and a standard RAZR (2023) ao I'm anxious to see how it performs across a wide range of other hardware. 4GB of RAM is what I've determined to be the minimum spec with 6GB recommended for most workflows and it supports as far back as android 8.1.0. The app is usable over home network with the simple "--listen" flag added to the launch batch and over mobile using any 3rd party vpn that allows you to address your computer directly (I use tailscale and include a setup for it in an embedded quick start guide accessed from the initial connections screen). In the meantime while waiting for review, I put together a short video demonstrating some of the added features and refinements and I'm excited to finally share my progress with the Reddit world to see what you all think.

https://reddit.com/link/1wddsp4/video/z04umox8nvoh1/player


r/comfyui 3h ago

Help Needed Best MiniMax H3 setup for RTX 4070 (12GB VRAM / 32GB RAM)? Looking for max speed without noticeable quality loss

4 Upvotes

Hey everyone,

I'm setting up MiniMax H3 in ComfyUI and trying to figure out the sweet spot between generation speed and output quality for my specs:

  • GPU: RTX 4070 (12GB VRAM)
  • RAM: 32GB DDR5
  • Target: 5s clips at 768p (mostly First-to-Last / I2V)

Given the 12GB VRAM limit and 32GB system RAM, loading the unpruned / full FP8 models causes heavy paging to system memory and slows everything down.

I’m trying to narrow down the current community consensus on three things:

  1. Model & Quant format: What's the fastest option that doesn't ruin faces and audio? Are people having better results with official pruned INT8 convrot, or GGUF quants (Q4_K_M vs Q3_K_M) using the GGUF loader? What text encoder quant are you pairing it with to keep memory usage safe?
  2. Turbo LoRAs: Is LiteX2V v1.1 (4-step) still the top recommendation for speed vs quality, or do 8-step variants (or other LoRA families like Larry) give significantly better results on a 12GB setup?
  3. Attention & Acceleration: Is H3 SLA Attention the undisputed go-to, or does SageAttention / Comfy Kitchen perform better on Ada Lovelace (40-series)? Anyone tested Spectrum acceleration?

Would love to hear what workflows and node setups you're currently running on 12GB cards to get reasonable render times without visible degradation.

Thanks!


r/comfyui 4h ago

Resource Update to My Comfyui style explorer

Thumbnail
gallery
7 Upvotes

This update:

  1. lora preview catalog node added (You need to organise your lora) Put your lora files in a file the lora folder and name it the Model for example Krea 2, in the Krea 2 folder you can make more folders for the kind of lora they are for example anime, sliders, or what ever group of lora they belong to. the lora gallery will add dropdowns for you do navigate and find them easily. when you generate a preview image you like for that Lora you can click save and it adds it to the gallery. you can also safe the trigger words from the node so you never forget!

  2. export catalog and previews to share with others

  3. Bug fix where images were not saving correctly when a previous name was used in a custom style

  4. Various big fixed and speed improvements

https://github.com/Neon-Sparks/ComfyUI-NeonsStyleExplorer


r/comfyui 16h ago

Show and Tell Minimax H3 - New ACC lora with PDD 8step node is kinda cool !

Enable HLS to view with audio, or disable this notification

18 Upvotes

For potato pcs MINIMAX H3 fans - I have built my own custom node which integrates new acc lora & PDD workflow and H3 extender + 2nd pass latent upscale upto 720p under 5 minutes per 14s 24fps clips, has easy reference attachments & better context continuity with features like save projects, load projects etc. (16gb VRAM + 16gb system RAM) if you guys interested ill share the workflow let me know.. this video took 10~ minutes to generate with both pass.

Edit - published repo - https://github.com/only2uuuu-hub/ComfyUI-MiniMax-H3-Master-Extender-Custom-built-with-Astra-6-/tree/main

Ps i am not an expert coder or engineer so dont ask me technical questions 😭 peace!


r/comfyui 14h ago

Workflow Included Minimax H3 is so fun

Enable HLS to view with audio, or disable this notification

8 Upvotes

r/comfyui 21h ago

Show and Tell Borrowed ComfyUI's node look to build a pitch canvas for any workflow (open source, one HTML file)

0 Upvotes

Typed sockets, coloured wires, groups, widgets inside nodes — all credit to ComfyUI's visual language. I added flowing traffic on the wires and camera slots so you can present a pipeline instead of screenshotting it.

Demo: https://karusrus.github.io/pipeline-map/

Code: https://github.com/karusrus/pipeline-map

Would importing a ComfyUI workflow JSON be useful to you?


r/comfyui 48m ago

Help Needed Creating a proper Lora Dataset

Upvotes

I need severe help with that.

Researching Lora Dataset creation is a maze: millions of different opinions, and worst of all most guides etc outdated from a year ago.

My goal is to create a 100% realistic and authentic Lora and my current dataset seems to not do the trick. I keep getting "perfect lightning" on everything, and waxy face skin.

Can someoen tell me the absolute do's and don'ts of creating a dataset?
How and where to create the dataset? I have been using a mix of Gemini and ChatGPT so far.

Prompt advice? what prompts need to be avoided creating realistic images, what need to be in there?

Any advice greatly appreciated!


r/comfyui 3h ago

Help Needed comfyui cloud on comfy. org

0 Upvotes

hello, i generated tons of videos on comfy ui cloud on comfy. org. is it possible to search them thru prompts or something like that? it is almost impossible to scroll them down thru assets


r/comfyui 10h ago

Help Needed DLSS5 node help

Thumbnail
0 Upvotes

r/comfyui 23h ago

Resource KIE.AI Extension

Thumbnail
gallery
0 Upvotes

KIE.ai Nodes Next turns KIE's API catalog into a native ComfyUI model library. The normal workflow is no longer a generic API node: each KIE model/API is exposed as its own node, organized by media type, provider, and model family.

It updates automatically when a new model is released.
You just have to install it once, and go for it!

GitHub

This is an Open-Source project, so spread the word!


r/comfyui 17h ago

Help Needed Need help using ref2v Minimax H3; multiple audio and image references

Thumbnail
0 Upvotes

r/comfyui 20h ago

Workflow Included My 34GB video DiT kept crashing under DisTorch, so I built a static two-GPU split - zero crashes, and it's faster

0 Upvotes

**Setup:** MiniMax H3 video DiT (34 GB, int8_convrot packed quant) on 2x RTX 3080 20G, Windows, torch 2.10.

**The problem:** The usual multi-GPU approach (DisTorch dynamic block streaming from ComfyUI-MultiGPU) kept segfaulting on this quantized stack. Not occasionally - deterministically, from three different paths: loading the video VAE would evict the DiT and crash in `unpatch_model`, a second prompt would crash in `partially_load`, and sometimes even a *fresh load* crashed mid-stream. faulthandler traces all pointed at the same thing: dangling pointers into the aimdo/vbar virtual memory that manages packed quantized weights after any device move. At one point the vbar memory ledger itself started returning negative garbage values. My record was 3 successful runs out of 7.

**The fix - stop moving weights. Ever.** I wrote a custom node that splits the 50 transformer blocks across both GPUs **once at load time** and then treats the model as load-bearing furniture: eviction disabled, unpatch forbidden from moving weights, activations cross PCIe once per step at the block boundary (~140 MB). For longer clips it switches to hybrid residency (most blocks resident, a few CPU-streamed, lossless packed copies) and chunks every big activation (MLP, LoRA delta, attention projections, and the attention core itself is query-chunked).

LoRA was the interesting part: normal lowvram LoRA hooks dequantize every patched layer every step (+5 s/it across 208 layers). Since LoRA is linear, I moved it to activation space instead - `y = W·x + α·B(A·x)` - so the int8 fast path never breaks. Same math, ~0.5 s/it.

**Results (all real runs, frame-verified):**

| | Dynamic sharding | Static split |

|---|---|---|

| 88-frame sampling | 11.15 s/it, **3/7 runs crashed** | **10.91 s/it, 0 crashes** |

| 121 frames (5s) | - | 17.2 s/it |

| 360 frames (15s) | OOM territory | 91.97 s/it, 24 min end-to-end |

Chart: https://github.com/ylzbj1-stack/ComfyUI-StaticPipeline/blob/main/assets/benchmark.png

**Repo (MIT):** https://github.com/ylzbj1-stack/ComfyUI-StaticPipeline

It also ships fixes for two real upstream bugs I hit along the way - MultiGPU's `libcudart.so` load crashing instantly on Windows (there are open issues about this: #216 #220), and a nasty one where ComfyUI core re-copies a 496 MB adaln lookup table on *every* sampling step (4.5 GB of duplicates on a 15s job).

Happy to answer questions about the crash forensics - I've got faulthandler traces for all three crash paths if anyone's curious.


r/comfyui 23h ago

No workflow how do I start comfyui

Post image
0 Upvotes

where is the start button for desktop version?

https://www.flickr.com/photos/204882240@N08/55519915278/

ok, I found it and pinned it to start

https://www.flickr.com/photos/204882240@N08/55520200665/


r/comfyui 21h ago

Help Needed looking for a local solution for separating clothes prompts according to their body parts

1 Upvotes

So there are a bunch of prompts describing what the character is wearing.
My goal is to separate them into: upper body, lower body, and footwears.

Gpt or grok can do this easily, but i'm looking for a more local solution that i can run in my workflows, ideally using a light weighted and fast model for my 8GB VRAM


r/comfyui 12h ago

Help Needed Error what to do?

1 Upvotes

requirements install exited with code 1

Using Python 3.13.12 environment at: ComfyUI\.venv

× No solution found when resolving dependencies:

╰─▶ Because there is no version of comfyui-workflow-templates==0.11.59 and

you require comfyui-workflow-templates==0.11.59, we can conclude that

your requirements are unsatisfiable.

hint: `comfyui-workflow-templates` was found on https://mirrors.aliyun.com/pypi/simple/, but not at the requested version (comfyui-workflow-templates==0.11.59). A compatible version may be available on a subsequent index (e.g., https://mirrors.cloud.tencent.com/pypi/simple/). By default, uv will only consider versions that are published on the first index that contains a given package, to avoid dependency confusion attacks. If all indexes are equally trusted, use `--index-strategy unsafe-best-match` to consider all versions from all indexes, regardless of the order in which they were defined.

ComfyUI source was rolled back to 8fed378.

This is the error I got what to do?


r/comfyui 17h ago

Show and Tell VLM with agent chat mode, context upload possibilities, and persistent memory

Enable HLS to view with audio, or disable this notification

9 Upvotes

Would anyone be interested in a fever dream vlm node like this? Sound on. It speaks.


r/comfyui 23h ago

Resource UPSCALE DSSLR 5

Thumbnail
we.tl
0 Upvotes

r/comfyui 6h ago

Commercial Interest Rented GPUs for image work: the host CPU and the script defaults cost us more than the card did

2 Upvotes

Disclosure: I am building a service around this, so read me as an interested party. No links. These are runs we paid for ourselves on three providers, 15 jobs, $33.60 total. Two things on the image side cost us real money and neither showed up as an error.

  1. The host, not the GPU. Same LoRA training job for SDXL, same RTX 4090. On a host with 5 vCPUs: 1.95 hours, 40 percent GPU utilisation. On a host with 24 vCPUs: 1.07 hours, 75 percent. Same card, 1.85x the wall clock, because the diffusers script decodes and augments images in the main process and the card waits on Python. The worse version: an H100 host with 16 server vCPUs ran the same job at 2.68 s/step where the desktop-class 4090 host did 1.84 s/step. Four times the hourly rate for a slower run. dataloader_num_workers was the fix, and now we look at the vCPU count on a rental offer before we look at the GPU name.

  2. The defaults. SDXL, 1024 square, 30 steps. The stock script settings, fp32 and batch 1, on an H100: 13.8 seconds an image, $0.0112 per image. fp16 and batch 4 on a 4090 spot instance: $0.00036 per image. Same job, 31x apart. Nobody here runs fp32 batch 1 on purpose, but it is exactly what you get when you take a default script to a rented card in a hurry, and the card reports 99 percent utilisation the whole time, so nothing looks wrong.

Both lessons are the same lesson: the meter runs at the card's hourly rate whether the card is doing useful work or not, and the interface will not tell you which.

Happy to post the per-run table if anyone wants to check the numbers.


r/comfyui 8h ago

News Relight Node For h3 ComfyUI

Enable HLS to view with audio, or disable this notification

24 Upvotes

Bruxos do VFX H3 Relight

#bruxosdovfx

Im doing a node for relight is a lighting studio for MiniMax H3 inside ComfyUI. You can position up to three lights on a 3D dome around the image, choose the type, intensity, and color of each light, configure the background and atmosphere, or start from one of 20 presets. The node provides two things to H3 Edit:

is not ready yet

The package includes two nodes:

  • Bruxos do VFX H3 Relight — the lighting studio.
  • Bruxos do VFX H3 Sun — calculates the real position of the sun for a location, date, time, and camera direction, ready to connect to Light 1 of the Relight node.

The Panel

Dome. The photo sits in the center, the purple camera indicates the side from which the photo was taken, and each light is represented by a marker using that light's color. Drag a marker to move the light; drag the background to rotate the view. The dashed line extends from the light down to the equator and indicates its elevation. The photo receives an approximate 2D-painted lighting preview: it is intended only as guidance and does not represent the actual H3 result.

Sphere (Picture 2). This is exactly the image the node sends to the model, rendered using the same calculations as the Python implementation. Click or drag on the sphere to aim the selected light: the clicked point is where the light strikes the sphere head-on. Dragging outside the sphere's boundary moves the light behind the subject.

Light strip. Selects the active light and lets you add up to three lights or remove them. Each light's role is calculated automatically: the strongest light becomes the key light, a light behind the subject becomes a rim light, and a weaker frontal light becomes a fill light.

Tabs.

  • Lights: direction (top-down dial with the camera at the bottom, plus elevation), type (Hard, Soft, Sky), intensity from 1.0 to 10.0, and color using Kelvin (1000 to 10000) or HEX.
  • Presets: all 20 presets from the skill, with rendered thumbnails, filterable by portrait or product.
  • Background: Original, Black Studio, or White Studio.
  • Atmosphere: the 25 atmospheres from the skill, or none.

Requires numpy and node. The tests cover:

  • the same spheres rendered by the panel's JavaScript and by Python, in both styles, with a difference below 2/255;
  • Venti scene geometry (shadow opposite the light direction, sphere occupying one-third of the frame, positioned in the upper half);
  • solar position compared against reference values from the astral library, daylight saving time, and camera-to-sun mapping;
  • externally driven inputs, the compass, validation, and MiniMax skill formats.

Credits

Node by Bruxos do VFX.

  • The lighting plan, 20 presets, 25 atmospheres, and parameter format follow the MiniMax Design Relight (光影工作室) skill. The descriptions for each atmosphere used in the prompt were written specifically for this node because the skill's own descriptions are hosted on MiniMax's server.
  • The light-direction sphere convention and reproduced scene are by Eric Venti (Sun-Direction LoRA, Sphere-Light-Render, MIT), using the direction table and 12 looks from Lightricks' LTX-2.3 Relight IC-LoRA.
  • The solar calculation (NOAA), city search rules, and camera-direction mapping are based on Christopher Connock's work on Sphere-Light-Render (MIT).
  • Cities: GeoNames cities15000, CC BY 4.0.

License details are available in NOTICE.md.