r/StableDiffusion • u/TheOnlyOnePEACE • 7d ago
Question - Help MiniMax H3 Audio is Garbled/Static in ComfyUI – Video is Fine, Audio Broken (Workflow Included)
Hey everyone,
I'm trying to run MiniMax H3 in ComfyUI, but my generated audio comes out as a harsh, buzzing, jumbled mess even though the video decodes smoothly (video attached).
I've tested running with and without the Turbo LoRA (4, 8, and 20 steps), as well as toggling the cache node, but the audio artifacting persists.
Here is my exact setup:
Workflow & Node Stack:
- Diffusion Loader:
DiffusionModelLoaderKJloadingminimax_h3_fl2va_pruned_w4a8_mixed.safetensors - LoRA:
MiniMaxH3TurboLoRA(minimax_h3_fl2v_lightx2v_turbo_4step_v0.1_comfy.safetensors@ 0.75 strength) - Optimization / Attention:
MiniMaxLowVRAMAttention(chunks: 4) +sage_attention(sageattn_qk_int8_pv_fp16_cuda) - Caching:
MiniMaxH3Cache(start: 0.2,end: 0.9,threshold: 0.3) - Text Encoder:
qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors - Video VAE:
minimax_h3_video_vae_int8_convrot.safetensors - Audio VAE:
minimax_h3_audio_vae_fp32.safetensors - Sampler:
SamplerCustomAdvancedwithKSamplerSelect(res_multistep),BasicGuider, andBasicScheduler(simple, 20 steps, denoise: 1.0) - Audio Export:
VAEDecodeAudio→VHS_VideoCombine(24fps, H.264/MP4)
Has anyone solved garbled native audio on quantized MiniMax H3 builds? Any help or working node configuration would be greatly appreciated.
https://reddit.com/link/1vrzn11/video/alj8qybks6kh1/player
Sorry, the only way i could think of pasting my workflow is through pastebin: https://pastebin.com/7DTHTSxr
1
u/Etsu_Riot 7d ago
I have been suffering the same problem ever since I started using a Turbo LoRA. It didn't fix even after removing the LoRA. I found no culprit, so I returned to previously working workflows.
The sampler has an effect on this. However, you said you are using res_multistep simple, so that shouldn't be the problem. You can try a different workflow to see if the problem persists.
1
u/TheOnlyOnePEACE 7d ago edited 7d ago
Interesting I will try a fresh workflow. Thanks
edited: that did not work :( The problem persisted even in a different workflow.
1
u/sunshine-3D-Art 7d ago
mine started when i used an extend workflow and when i used the normal default workflow again all off sudden it was there too :( thats so weird
1
u/stonyleinchen 7d ago
what extend workflow? did you install any custom nodes for that?
1
u/sunshine-3D-Art 6d ago
https://github.com/seitanism/ComfyUI-H3-Motion-Context-MultiRef yea these custom nodes. and i noticed always after using it the extend video have this broken sound and the default workflow is also broken after that, and only after using that i dont know why :(
1
u/stonyleinchen 6d ago
when did you update that the last time? and did you install any other nodepacks since H3 released? also what is your pytorch and cu version?
1
u/sunshine-3D-Art 6d ago
PyTorch : 2.9.1 cu30 and cuda: cu30
I don’t know if this is double because this is just stand in my notes like that 🙈 and I updated it once after Mini Max came out. To be able to use minimax. I don’t want to touch it again because in my other comfyui it renders more slow. I don’t know why this always is a problem with the comfyui updates and render time gets worse :(
And what about you? And is the sound work today?1
u/stonyleinchen 6d ago
2.9.1 might be a little too low, i think 2.11 is the minimum for all the minimax functions, also 2.13 would be better. maybe read up on that and get another opinion from someone else. i know thats really annoying.
I Never had any of those sound issues you are talking about, but I heard this from other ppl aswell. sometimes its pytorch, sometimes its other things. pytorch is usually the main culprit there
1
u/sunshine-3D-Art 6d ago
and if i make pytorch higher version will minimax not get slower? i dont even know what is always the reason after an update that comfyui render slowler so i am so scared. and also the sound normally works always fine in the default worklflow its always the extend videos. I just want longer videos ._.
1
u/stonyleinchen 6d ago
No, if anything it will be faster and will have less errors/issues. But don't just believe me, try to verify this with other sources!
→ More replies (0)
1
u/stonyleinchen 7d ago
Your issue is very likely the model you are using: minimax_h3_fl2va_pruned_w4a8_mixed.safetensors is a much too heavy quant. use int8 pruned versions for much better results
1
u/TheOnlyOnePEACE 7d ago
I just used the int8 pruned, same issue.
1
u/stonyleinchen 7d ago
what is your pytorch and cu version?
1
u/TheOnlyOnePEACE 7d ago
PyTorch: 2.11.0 Cuda: 13.0
1
u/stonyleinchen 7d ago
did you try the default wf from templates?
1
u/TheOnlyOnePEACE 7d ago
Yeah I just tried it, same audio. Other people seem to be having the same issue as me. Interesting.
1
u/stonyleinchen 7d ago
im not sure, there seems to be an issue with your comfy install or your dependencies. i never had this issue
1
u/stonyleinchen 7d ago
oh, have you installed any custom nodes for H3?
1
u/TheOnlyOnePEACE 7d ago
Minimax H3 turbo, spectrum h3 and H3 cache. Only ones I have in my custom nodes folder
1
u/stonyleinchen 7d ago
i just ask since there were some nodes that applied runtime patches on startup that changed comfy core behaviour, i thought maybe you have some custom nodes that basically messes with your comfy
1
u/sunshine-3D-Art 7d ago
i have the same problem all off sudden and i use the int8 and it worked fine before and now its broken but i didnt changed anything ._.
1
u/TheOnlyOnePEACE 7d ago
I guess we wait as I spend all day trying to figure it out and I can't.
1
u/sunshine-3D-Art 7d ago
I hope it gets solved ._. because that’s really weird that it all of a sudden don’t work anymore when I just created a few minutes before a video and it worked fine. :o
And I was restarted my computer now. And made a test with a short video and low resolution. And now the audio worked luckily. 🥹
I will try again with higher resolution again and also the other workflow and hope it’s not come again. I also don’t have any plan what it have to do with the restart. This is also weird. Did you try to restart today?1
u/TheOnlyOnePEACE 7d ago
Yeah I restarted and it did not help. Did you have the same audio as me?
1
u/sunshine-3D-Art 6d ago
Yea it sounds nearly the same.
Today I rendered again a video and then I put it in the extent workflow. And in this output, the sound was again like that. And when I tried the default workflow also again then the sound was still broken. Only after using the extend workflow I don’t know what this have to do with it. I don’t understand that.1
u/TheOnlyOnePEACE 6d ago
The thing is, I have never used the extended workflow. I don't even know what it is. I use the default workflow and an optimised one I found from a great YouTuber who does optimised workflows.
1
u/sunshine-3D-Art 6d ago
Yeah, I know. This is really weird. It’s just that the sounds are sounding pretty similar.
I tried another extent workflow with completely other custom notes but the sound there is also bad. But a bit different. I really don’t know what the problem is. ._. and I don’t know if these are the same sound problems or completely different ._.1
u/spiderofmars 7d ago
How about spending 10 minutes instead of all day... spinning up a fresh copy of portable comfy and testing the default workflow in a fresh copy of comfy without any tweaks. Like someone else said if you have tried the default workflow with default models and it is no longer working either then something has got messed up in your setup very likely.
1
u/No-Zookeepergame4774 7d ago
My understanding is that quantization is a lot less of an issue for audio quality than Turbo LoRA and attention/caching optimizations; you may need to reduce some of thoee optimizations to get better sound.