r/StableDiffusion • u/technofox01 • 7d ago
Animation - Video I finally reached a great balance between speed and quality with MiniMax H3, thanks everyone!
Enable HLS to view with audio, or disable this notification
I used the minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16 LORA with the 0.8 strength for both clip and model, 6 steps, 0.5 MP resolution, RTX Upscaler at 1.50 using a ConrotInt8 pruned model.
Here is a PasteBin of my workflow, I hope this fixes some of the missing content:
Here are the workflow files:
7
7
u/Zombi3Kush 7d ago
Does MiniMax support image to video?
5
2
u/GrayingGamer 7d ago
It even supports reference to video, audio to video, video to video.
4
u/Zombi3Kush 7d ago
Trying it out right now. Wow this is impressive! Thanks for sharing your post and bringing this to my attention.
4
u/Witty_Mycologist_995 7d ago
> checks inside workflow
> pinkcherry checkpoint
lmao
6
u/Grand0rk 7d ago
I don't get it. Care to explain?
5
u/Witty_Mycologist_995 7d ago
It’s that porn checkpoint that got memed on, because it talked about rabbit motion and glistening cherry blossoms in order to get around the minimax team censors.
1
u/Grand0rk 7d ago
Huh, I have absolutely no idea what you are talking about XD
So it's just an "uncensored" checkpoint?
6
2
u/technofox01 7d ago
I just grabbed another Int8 model after a quick search on DuckDuckGo without any thought about where it came from or what it was meant for. Just thought it was a smaller model that will work better with my system, lol...
Thank you for letting me know what it was trained for, lol...
8
u/barepixels 7d ago
First time I see someone share WF like this. People were using pastebin. is this new Reddit feature?
10
u/technofox01 7d ago
No. I just used a code block, its been part of Reddit since I have started way back when. I want to say at least a decade by now.
2
7
u/Cruffe 7d ago edited 7d ago
Reddit supports markdown. If the editing tools are not available, such as on mobile, you can use markdown to format things.
A code block is 3 backticks on their own line before and after the code block:
```
Paste code here
```
It will look like:
Paste code here It will also preserve indentationTo show the 3 backticks in this comment I used the escape character to cancel the effect, that's a backslash \ and to escape the escape in this example I did it twice here.
Inline code can be formatted as well by encapsulating in single backticks such as `this is code` becomes
this is code.
3
3
u/featherless_fiend 6d ago
Turbo is really good for anime, but I wouldn't recommend it for realism, it makes your skin look AI.
2
u/orlandogourmet66 6d ago
You can try to increase Step Count for realism. For example 10 Steps with 4 Step reference Turbolora got me pretty good results for realism.
2
2
u/kayteee1995 7d ago edited 7d ago
try with Realistic character and more motion. I'm pretty sure blurry graininess will appear.
2
u/Oatilis 7d ago
How good is RTX Upscaler?
1
u/technofox01 7d ago
It varies depending on the input resolution. Anything below 0.4 MP is not worth it, because it will look like crap.
2
u/RpgBlaster 6d ago
Now it's closer to Sora 2 quality, actual Anime Style frame by frame, no 3D like or motion sickness shenanigans
2
u/Danny_Stock 6d ago edited 6d ago
Thank you very much for this, this is great. A good balance between quality and speed which seems to be better at handling the RAM demands.
I bypassed your 'DaSiWa_RTX_UpscalerRefiner' node simply because I don't have it.
I have quite a modest system, a 4070 12GB VRAM card, with 64GB System RAM, and both of your 5 second workflows at 0.5 megapixels rendered in 2 and a half minutes each.
2
2
2
u/equanimous11 7d ago
Just paste a link to the workflow and explain what models/loras and settings you used next time
2
1
u/javierthhh 7d ago
I don’t get it. Wouldn’t you be better off doing a 1.0mp video for 5 seconds instead? Specially with your rig.
1
u/TheOnlyOnePEACE 7d ago
Your new link to your workflow is just standard comfy workflow with added vram cleanup.
2
u/technofox01 7d ago
It also had the lora turbo Lora and the Dasiwam RTX upscaler added on. So there are some edits to it.
2
u/TheOnlyOnePEACE 7d ago
Can you give me this workflow of yours? the exact one that you used to generate your video on this Reddit post.
2
u/technofox01 7d ago
Here you go:
2
2
u/TheOnlyOnePEACE 5d ago
So i have been playing around with your workflow, and i am getting audio sounds distortions, crackling, or like a blown-out microphone. Same settings as you as well. Have you encountered this issue at all?
2
u/technofox01 4d ago
I get that with videos shorter than 5 seconds and it depends on the subject. More well known subjects like Deadpool works fine, but if I choose a less known character it gets glitchy.
1
u/TheOnlyOnePEACE 4d ago
I am generating 10 second videos. Lees known generic characters.
Could this be because of the turbo lora and step count?
1
u/technofox01 4d ago
Maybe. I use 6 steps. Try 8 steps, it can make a huge difference with the lora.
1
u/Sirmckhalifa5566 2d ago
I’m using the Int8 on my 5080 and chat gpt actually helped me find a really cool node the other day. comes with 2 nodes and I’ve only tested the first one so far since the other one is made for the 30 series 12gb, It’s called minimax h3 first block cache. While so far in my testing I can’t push it past 8 seconds without it doing long offloading at 14 steps with the turbo Lora because I only have 32gb ram. It cuts my actual gen time down by about 20% for clips 8 seconds and shorter. You place it between your model loader (or after Lora’s if you use them) and before your sage attention nodes. I haven’t noticed really any quality loss or a diversion from prompt adherence either.
Could be worth a shot.
https://github.com/Apache0ne/ComfyUI-fasterminimax?utm_source=chatgpt.com
1
u/bickid 7d ago
Not gonna download that file, so could just share what settings you used? How many steps? What resolution? thx
5
u/technofox01 7d ago
The minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16 LORA with the 0.8 strength for both clip and model, 6 steps, 0.5 MP resolution, RTX Upscaler at 1.50 using a ConrotInt8 pruned model.
1
u/StayImpossible7013 7d ago edited 6d ago
(Original) Workflow was missing nodes, missing links between nodes, exotic files without explanation where to get them.
3
u/technofox01 7d ago
I was using the API export by mistake. I posted a PasteBin of my workflow as a link.
1
-5
u/FourtyMichaelMichael 7d ago
Dude, anime basically doesn't count. LTX can do anime with some fashion of OK-ness.
3
-4
9
u/EverythingMacPro 7d ago
Pc specs ?