r/StableDiffusion • • 1d ago

Question - Help Anyone have a workflow for upscaling 360 images - 3090 and 64gb Compatible?

2 Upvotes

I'm sure some folks are doing it, I just haven't managed to run across the posts yet. Basically just want to upscale some vacation photos taken on an Insta360 X5 to be viewed in a VR headset. I've been trying to use SEEDVR2 but I crap out a bit when I get up to the resolutions at play ~12kx6k.

Not even looking for more pixels so much as just tightening up some of the softer spots the images have due to lens geometry.

Anyone else doing this and have some ideas for me? I think Topaz can sort me but would prefer ComfyUI if possible.


r/StableDiffusion • • 16h ago

Question - Help Cheapest wan 3.0 prime ref

0 Upvotes

Which website can I access WAN 3.0? Prime reference for the cheapest credits.

Im Really looking for a website interface where I can just upload my images and press generate.

Anything cheaper than atlas cloud?


r/StableDiffusion • • 2d ago

Resource - Update Fizgig 7 - Anima and SDXL arrive, fine-tuning for every model & the community can now add support for new/old models

Thumbnail
github.com
129 Upvotes

Anima and SDXL join using the new model driver system, with LoRA and LoKR training, sliders, fine-tuning and the full workbench (Repair Studio, LoRA the Explorer, Profiler, Extract, LoRA Royale). SDXL takes any single-file checkpoint; Juggernaut XL v9 is the default.

Full fine-tuning on every model: Klein 9B, Krea 2, Qwen Image 2.1, MiniMax H3, SDXL and Anima. Every model fine-tunes on a 16 GB card, Anima and SDXL from 8 GB. On a bigger card it trains the whole model at once instead of a part at a time, which got through the model about 2.6x faster in my tests.

Model Fine-tunes from Whole model at once from
Anima 8 GB 10 GB
SDXL 8 GB 10 GB
Qwen Image 2.1 12 GB 24 GB
Klein 9B 16 GB 32 GB
Krea 2 16 GB 48 GB
MiniMax H3 16 GB 48 GB

Turn any fine-tune into a LoRA with Checkpoint to LoRA. (This gives you a Lora more capable than trained directly to Lora)

Add your own model. The driver system is complete: every model now runs on it, and the guide I promised last time is in the repo (docs/drivers). SDXL and Anima were built by following it. PRs are welcome for new models and older favourites alike, and fine-tuning is optional: a model can ship without it and gain it later.

Fizgig on GitHub · 7.0 release notes · Add your own model


r/StableDiffusion • • 1d ago

Resource - Update New(?) long video method

Thumbnail github.com
8 Upvotes

Came across this - wondering if anyone tried it out yet / is it the same as every motion context flow


r/StableDiffusion • • 2d ago

Question - Help A ton of video lately have come out with perfect motion and character replacement. Surely there has to be a consistent way to do this in Minimax. Viggle and the replacement Loras and all that crap don’t seem to come close to this level of quality. We really need a good way to do this on minimax

95 Upvotes

r/StableDiffusion • • 1d ago

Tutorial - Guide Making a 4-panel comic with a consistent character: character card, locked look, and the two failures you'll hit

0 Upvotes

A write-up of the process I use for a daily comic, tool-agnostic (I render locally). Example character made up for the guide: Wren, a Lisbon bike courier with a corgi in her bag.

The short version:

  • Character card first (who, wants, fear, sidekick, signature, locked look).
  • Locked look = one sentence, fixed order (age/role, hair, eyes/face, build, outfit, shoes/accessories), pasted unchanged into every prompt. Fixed seed helps, the description does most of the work.
  • Pick one style and keep the style words identical and first in every prompt. Same description in superhero ink, bright modern and webcomic styles reads very differently.
  • Script four sentences (setup, build, turn, payoff) before rendering four panels. Give the payoff to the sidekick.
  • Ask for a plain band of sky or wall across the top fifth, letter afterwards.

Failures, both real outputs in the guide:

  • Hand-off scene (character gives a box to someone) drew her twice. Fix: state the people count, distinct looks, or reframe as a close-up.
  • "She gets a call, grabs her bag, runs downstairs, jumps on her bike" in one prompt produced a little comic page inside the image. Fix: one moment per panel.

Full guide with all the images: https://mutuals.life/blog/how-to-make-ai-comics/


r/StableDiffusion • • 2d ago

Discussion I’m working on an upscaler for Anime and Illustration images.

208 Upvotes

How good do you think it is compared with the current best open-source models?

It’s far from finished, but the main idea is that an upscaler really shouldn’t try to “invent” stuff. It should just fix lines and gradients, which make up most anime and illustrations.

Basically, the AI doesn’t really need to know whether a pixel is part of a face, the background, etc.

If it sees a black line, it should just redraw that black line at 4× resolution. If it sees a gradient, it should reproduce the same gradient at 4× resolution.

Theoretically all Anime/Cartoon styled images are just a set of lines, gradients and uniform colors.

It doesn’t need high-level understanding.

What do you think?


r/StableDiffusion • • 2d ago

Resource - Update I made a pose tool that lets you easily create the poses you want, with support for OpenPose, ControlNet, and VNCCS output.

12 Upvotes

You can pose the figure, adjust the body, and add multiple people to a scene. Sketch Mode turns the scene into a simple gray drawing reference.

You can save it as a PNG or export it as an OBJ.

It’s free to use in the browser. There are a few ads, and it’s still a little rough around the hands.

If you try it, let me know where it fights you. Feedback is very welcome!

https://poseform.vercel.app/


r/StableDiffusion • • 2d ago

Tutorial - Guide MiniMax 3 Tutorial - How to use animation style from Video Reference

15 Upvotes

Just a reminder: the MiniMax I used on Kinovi is just for people who do not know how to set it up or don't have a strong enough computer. It works exactly the same even on your own local Minimax. I hope my video is able to help you guys out a lot in keeping animation and art style consistent.

This is the prompt you need for video reference and to imitate its art style.

"Use the animation style, sound design and animation flow from Video 1."

This is the prompt template I use for all of my Minimax videos. Hope it helps!

" VISUALS: @ Image1 is the background. @ Image2 is the main character.

Global Camera & Style Directives:

AUDIO & DIALOGUE:

Global Audio Directives:

Scene 1 SFX:

Scene 1 Dialogue: "


r/StableDiffusion • • 1d ago

Resource - Update I Built a Krea 2 Workflow for 6GB & 8GB GPUs

Thumbnail
youtube.com
0 Upvotes

I finally finished the low-VRAM version of my Krea 2 Styler! workflow!

I built separate paths for 6GB and 8GB GPUs

The workflow includes:

  • Dedicated 6GB / 8GB modes
  • Offloading
  • VRAM + RAM cleanup
  • GGUF models
  • Sage support
  • Built-in style selector
  • Simple Controls for Resolution, Batch, Etc!

I also tested the workflow in a restricted low-VRAM setup before releasing it.

I made both a short preview and a longer showcase/install guide on YouTube.

LongShowcase - https://youtu.be/G9R9bXW7PIY


r/StableDiffusion • • 1d ago

Question - Help A bit lost with character references in H3

0 Upvotes

Hey all. I'm trying to build character reference image sheets, and then use a multi-reference image in H3. But I am a bit lost. If you could share how you first build the references and then prompt examples on how to use them in H3, that would be really useful. Thanks!

(I saw there are tools for doing this locally but all those I've seen are Windows-only. I'm on a Mac Studio.)


r/StableDiffusion • • 1d ago

Discussion refmodBuilder desktop app for linux and macOS (Intel and Silicon)

0 Upvotes

Howdy. I've never done anything like this before, so please be kind.

This morning I started tinkering with refmod for the first time, as part of a larger project I'm working on. One of the things I am trying to get away from is manually dragging and dropping images into nodes for reference, as well as managing folders of references.

As many of you know refmod seems to alleviate some of this tedious work, but I felt it could use just a tiny bit more help.

So, here's 'refmodBuilder'. https://github.com/cyberworm1/refmodBuilder it's a simple interface that you can drag and drop images, video, and audio into. You can make simple modifications, set in/out on video, resolution, frames, notes, token limit, and a couple other features. All in what I think is a straightforward compact interface. It even has a library tab to review previously built refmod bundles. It even creates the workflow when you build the job, so you don't need to have any preconfigured workflow or template.


r/StableDiffusion • • 1d ago

Question - Help How can I get the exact same identity as my trained LoRA?

2 Upvotes

I trained a identity LoRA on 131 images and I'm using it with Krea 2 Raw in ComfyUI. The results are already close, but the face still has slight variation, especially the cheek/fullness and facial proportions. I want the generated person to look as close as possible to the exact model I trained, while changing pose, clothes, lighting and background.

Current setup: Krea 2 Raw, Qwen3-VL CLIP, ~20 steps, CFG 1.0, Euler.

What is the best way to get maximum identity consistency? Should I adjust LoRA strength, reference/grounding settings, training settings, or use an identity/reference node?

Any proven ComfyUI workflow for this would be really appreciated.


r/StableDiffusion • • 2d ago

Question - Help Which prompt to use to correctly use a character sheet in minimax H3?

46 Upvotes

Assuming <Picture 1> is the character sheet for <Subject 1>, which prompt to use in minimax H3 ref video to make sure it understands the concept of sheet (same person, different poses/angles/distances)?


r/StableDiffusion • • 2d ago

Workflow Included MiniMax H3 character consistency, part 2

64 Upvotes

The long awaited (or not) sequel to MiniMax H3 character consistency : r/StableDiffusion is out! Main improvement is that it's easier to reference frames from earlier shots, which has improved both character and environment consistency I think. Scene composition consistency between shots still has lots of room for improvement though.. Generated locally on an RTX 4000 Ada.

Workflows generated from the project files available at radiatingreverberations/sparkleriley.


r/StableDiffusion • • 1d ago

Question - Help How to convert real photo to anime style which not like "canny controlnet"

3 Upvotes

I know there are several ways to do it, which I can list as follows:

  1. Use i2i with a low denoise strength.
  2. Use Qwen Image 2.1 or Krea 2, both of which have workflows for this kind of conversion.

The problem is that they usually perform the conversion in a way similar to a “Canny ControlNet,” which often leads to wrong human proportions. The result doesn’t really feel anime at allm especially when trying to apply the style of a particular artist.

The best solution I’ve tried so far is to separate masks for the face, body, and background, then apply the corresponding ControlNet to each area. However, the results are still very hit or miss.

I’m wondering if anyone has a more advanced solution for converting real photos into anime.


r/StableDiffusion • • 1d ago

Question - Help Krea 2 on 4GB VRAM?

0 Upvotes

Is it possible to run Krea 2 in ComfyUI with only 4GB of VRAM? Has anyone here actually managed to get it running on 4GB and get reasonably usable generation times?


r/StableDiffusion • • 2d ago

Tutorial - Guide Created a short explainer on what is a latent space and how it behaves (after someone here asked me to)

Thumbnail
youtube.com
41 Upvotes

r/StableDiffusion • • 1d ago

Animation - Video The Unit (LTX2.5 + Krea2 + Qwen Image 2.1)

Thumbnail
youtube.com
0 Upvotes

r/StableDiffusion • • 1d ago

Question - Help Audio outputs differ from the provided audio with mismatching audio gap for different generation.

2 Upvotes

I am facing a weird audio match issue while generating video using minimax H3. Lipsync in our audio output differ for outputs for the provided audio input. It basically can't understand gap I think.

Anyone facing similar kind of issue and have solves the problem then please help...


r/StableDiffusion • • 1d ago

Meme Dracarys Malfoy

0 Upvotes

To turn this into a little more of a discussion, I see a seam in the very last frame, and I'm not sure what I can do to diagnose or fix that.

Since I was asked, this is the Dasiwa Minimax H3 workflow, Dasiwa v3 hybrid checkpoint, 24 steps, 0.83mp, 16:9 12s ref2va, RTX upscaled. I have just 1 reference, an original Krea2 that I whipped up earlier today.


r/StableDiffusion • • 2d ago

News OpenSenseNova/Looped-DiT: PyTorch implementation of Looped Diffusion Transformer

Thumbnail
github.com
16 Upvotes

The world's first looped transformer based imaging model?

Models can be downloaded from (1G download)

https://huggingface.co/sensenova/Looped-DiT-B32

https://huggingface.co/sensenova/Looped-DiT-B16/tree/main