“A lot of animation styles”. Support merged to Comfy official.
Weights
text encoder … qwen 32b - even with 4bit its 20Gb - available in nvfp4 as qwen3vl_32b_minimax_h3_nvfp4_awq
nown issues at this point:
- audio breaks with stochastic samplers (SDE/Ancestral)
- audio probably breaks with inpainting and when using denoise under 1.0
int8 error is so low that it shouldn’t be huge difference at all … twice as fast and only ~0.9% relative quant error [compared to fp16]
it has it’s own streams for image/video/audio, you can mix and match them.. up to 9 images, 3 videos, 3 audio
it can be slower if you use a large image, the image size affects it
The old man from <Picture 1> is sitting in a studio in front of a blank canvas, he paints the <Picture 2>.
Works fine with RES4LYF samplers, though you need to update RES4LYF to most recent version (from today) to get around a small error with nested tensors. Also you really do have to set SDE noise to 0.0 right now (eta in ClownSamplers), it totally breaks the audio otherwise.
GH:xmarre/ComfyUI-Spectrum-MiniMax-H3 accelerator node “Is this better than Easy Cache or can you use them together”
patientx: “int8 model + nvfp4 clip + sage-attention + spectrum”