My ComfyUI workflows for using MiniMax H3
This workflows are used by me to create my art.
They are optimized for my checkpoints and created of my latest knowledge to enhance the outcome.

Versions & Information👇👇👇👇👇👇👇👇
👉 Please read below and the file descriptions "About this version" for more info's.
⚠️ Do not use the workflows with the "Nodes 2.0 beta" from ComfyUi or it will mess up things.
Types of workflows
MythicAlchemy C-MMH3
🎬 All MiniMax H3 modes - T2VA, I2VA, FLF2VA, REF2VA in one workflow
🖼️ Reference Director - drag-and-drop media, slot-based ordering, trims, per-media prompts
🎵 Audio support - embedded video audio (V/A/V+A switch), standalone audio tracks
🎯 Smart resolution - automatic aspect-ratio fitting and megapixel presets
⚡ RTX upscaling - optional real-time super-resolution (RTX GPUs)
💧 Watermark overlay - custom branding on final output
🚀 Performance optimizations
🃏 Last Frame Extraction

🩻 Known issues and advice's
Install ffmpeg!
Update Comfyui and custom_nodes!
Make sure to read where files/models should be placed inside the workflow
Check if the filepath for model/clip/vae match your system like Linux/Windows
The plugin ComfyUI-DD-Translation can break node connection (avoid)
Clone the Repo's from GitHub if something is missing like newest Kijai Nodes
All older Versions are available inside my GitHub Repo.
YOU are responsible for outputs as always! If you make ToS violating content and I get aware I WILL report this.
Description
Updated
FAQ
Comments (161)
YEEEEESSSS YEEEESSSS OMG MAN YOU ARE THE BEST!!! I can't wait to test it thank you so much!!! ♥(ノ´∀`)
DaSiWa is on vacation and STILL drops a day one workflow for MiniMax H3 hahaha, are you human or some sort of machine???
Hahaha, yeah still on vacation 🫣
@Darksidewalker enjoy your break, thanks so much for this workflow!
Legend 🫡
literally just dropped my quick Ref Workflow right before this, i was waiting for a good WF to come and did not expect one of my fave wf creators to drop on day one! amazing thank you :)
Yes! Now I can melt my GPU with the most efficient workflow! 😁
Thank you! Melt that thing!
Much much appreciated. You da man!
Base model can now be set to MiniMax H3 for better visibility of this workflow.
Thanks, was not possible yesterday!
@Darksidewalker I figured.
Awesome! What package does the MiniMaxH3MemoryEfficientSageAttentionPatch node live under? (i updated comfyui and all my packages but it's still unrecognized.)
Same issue here!
you need to update KJNodes: https://github.com/kijai/ComfyUI-KJNodes
ComfyUI wont say it needs an update sometimes, so manually git it into the customnodes folder
@mrweaz ComfyUI-KJNodes updated, still MiniMaxH3MemoryEfficientSageAttentionPatch node is missing
@myprivacy27091991221 Same.
@hardwire666 Yeah I’m having problems with the director and director guide nodes
@myprivacy27091991221 yeah, i don't now how to fix that, i am having the same issue
@myprivacy27091991221 in that case I'm guessing you don't have sage-attention installed. Try installing that and see if it works.
@mrweaz thanks it was a couple of things missing on my end, had to use chatgpt to walk me thro cuda tourch and a bunch of other updates... i hope it didn't break my setup for something else lol
checkpoint? lol. i cant wait
Will need time for this to brew 😸
You are god. YYDS.
we need a way to do character swap. It's about the only thing that's not consistent with this model.
I'm somehow sure (not tested) it should be possible with ref2va mode
Looking forward to see your future merges of this model.
looks great but it also needs https://github.com/DemonGatanjieu/Anomalous_Model_Browser?
only hardcore CivitAI users use that browser and shove it down everyone’s throats, causing missing node errors to appear. I wish people would make separate WF without using that abomination of a browser.
@blhll I do not plan to implement this ...
Fantastic workflow as always. Just one problem. Does your Video Combine node filename_prefix has no number counter? It keep overwriting the previous generated video file since the filename number stuck at _00001.
Its used with time placeholder so technically it cannot overwrite, but I can look into this when not using the placeholder ✌️
@Darksidewalker Tested your updated node pack. I have to say, your node pack are damn amazing. Still, while I can use time placeholder, would still very much prefer to save the video with simple filename with sequential numbers for specific cases like 'XXX_00001, XXX_00002' like the comfyui default one.
Im a newbie and wonder if anyone know what to do i ask ai but nothing helps. I try to run with Sage with this node and get this error: \ltxv_nodes.py", line 2034, in execute
raise RuntimeError("sageattention is not new enough version or could not determine CUDA architecture, cannot apply MiniMax H3 Memory Efficient Sage Attention Patch.")
RuntimeError: sageattention is not new enough version or could not determine CUDA architecture, cannot apply MiniMax H3 Memory Efficient Sage Attention Patch.
As described in the Known Issues section: Update KJNodes from GitHub https://github.com/kijai/ComfyUI-KJNodes
This is the challenge. I have done this. :/
@johansen777 Maybe you did not install sage-attention.
https://www.youtube.com/watch?v=peq4TFOxHW4
This video helped me just now when I ran into your same issue johansen. Verify your cuda, python, and torch version. Find the corresponding triton you need (uninstall and reinstall if you have an older version), and then find the correct sage attention wheel for your setup. Good luck!
Easiest way is to use DaSiWa's comfyui installer for a fresh portable install then just copy over your models folder. it auto-installs everything you need including triton and sageattention. Wish I had known about it way sooner.
Have you tested Spectrum vs. Easy Cache? https://www.reddit.com/r/StableDiffusion/comments/1vf1ze3/spectrum_acceleration_for_minimax_h3_in_comfyui/
There's some chatter about Spectrum degradation being far less than when using easycache. Curious as to your thoughts!
ive been testing everything that has been suggested, SageAttn, Easy Cache, Sol-Attn and Spectrum. so far the only method that has given a good speed boost is Spectrum, the others are very minor (seconds) with a loss of quality that doesn't make up for the time you save. Spectrum is the only one worth using so far. ive replaced the EasyCache and SageAttn in the wf and it gens alot faster with minimal quality loss, i suggest everyone to do the same.
Things are going fast and I just released day1. There might be huge updates. I will also make a HUGE update on my director.
Thank you for mentioning it!
I'll look into! ~ The WF already has a speed improvement like 1.2~1.61 compared to the basic one.
After some testing, the acceleration methods like MemCache, Spectrum are really lossy. Only SageAttention is near lossless. Deforming, motion alteration and ghosting can happen.
@Darksidewalker This is already awesome. Can't wait for the next update!
@Darksidewalker I had the opposite results for some reason, Spectrum was the best result I had gotten. Maybe it's hardware specific? I'll run some more tests with Sage and EasyCache and re test 😂
Tested Spectrum vs MemCache vs Sage-Only vs No-Speed-Up ~
4s, 0.65MP, Fixed Seed, 4:3, I2VA ~
MemCache + Sage 56s
Spectrum + Sage 80s
Sage-Only 107s
No-Speed-Up 173s
I posted a side by side comparison on my discord.
@Darksidewalker ill have to have another go, i may have been setting something up wrong. thanks for the info :)
Like usual, you're the king! ♥
MiniMaxH3MemoryEfficientSageAttentionPatch node is missing, even after updated al nodes
they're on ComfyUI KJNodes Nightly :)
@Ertyer Still doesn't work. Even with a manual install using the command line.
You have to git clone, since he did not update the version
@Darksidewalker I can't speak for OP but git clone didn't do it for me. I had to drop to torch 2.11 + CUDA 13 and then manually run
python -m pip install --upgrade "triton-windows>=3.6,<3.7
and
python -m pip install --force-reinstall --no-deps "https://huggingface.co/ussoewwin/Sage-Attention-for-Windows/resolve/main/sageattention-2.2.0%2Bcu130torch2.11.0-cp313-cp313-win_amd64.whl"
Seems to be working now.
git clone helped me. Switching to latest/nightly did not help.
@Ertyer thanks
i'm having trouble with the audio only; I've followed the official prompting guide using the: She says: <d>[English] Bla bla bla</d>. and overall_soundscape: sounds
but I keep getting weird sounds and no dialogue.
I tried inputing the prompt in the media prompt section, then in the global prompt section but no good results. Any tips on what am I doing wrong?
Just to add onto this, I have been getting garbled/cut off chinese audio at the end of the generation versus the official workflow.
Was a bug in the frame -> duration calculation, should work now, please update the node pack :)
@Darksidewalker Thank you! Love your work as always btw! ❤️
@Darksidewalker I still get extreme harsh sounds instead of anything else :(
@KiraNugget Did you try without memcache and sage? maybe prompting?
@Darksidewalker I figured what was hapenning, when installing the workflow I saved it before installing all nodes and for some reason the audio shift was at 0.01 instead of 4.00. It works like a charm now.
just to make sure, shift video should be at 11 and audio at 4 right?
@KiraNugget I tried 11 and 4, but 12 and 3 are default
I got good results with 10 and 5, too... still testing...
There was a problem with this; change the scheduler to Simple. You probably have Normal or Beta set up, and I have issues with sound when I use those.
Great WF but I am not sure how to use the ref2v part. Do I just add all references to the Director timeline and them somehow mention them in the prompt?
haven't tried ref2vid yet but here is the official prompting guide https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.md
Yeah you drop them in, and prompt for <Subject 1> is the person in <Image 1> wearing a fantasy tunic with black hair. Then you would reference <Subject 1> in your prompt following the prompt guide like so:
integrated_multimodal_description: [Shot 1] <Subject 1> is standing in a forest filled with ash, he looks around as the camera pans to the left..
overall_soundscape: ...
non_diegetic_music: ...
@mrweaz cool but how we will see or tag the image? (for multiple images)
@Keithxx you do the same for <Subject 2> etc as long as you tell the model what each subject is it doesn't matter so much which order you place them in. But if you do <Image 1> <Image 2> etc then you need to remember what order you imported them in.
@mrweaz hm i did uploaded all at once, but ill use that way
I was able to use up to 9 images at once in this workflow.
@Acleveralias figured it out it was a ui bug, i changed the mode back and forth, fl then ref and it worked
If you add a video in V+A mode in the video lane as well as an audio in the audio lane, which of them is <Audio 1> ?
@maway the prompting guide on Huggingface states this:
Visual and Audio Tracks from the Same Reference Video
<Video N> and <Audio N> are numbered independently. Each index indicates only the label's order within its own category and does not encode a pairing between the two categories. The same reference video may therefore correspond to <Video 1> and <Audio 2>; different indices do not prevent them from coming from the same source asset.
An ordinary reference video does not create <Audio N> merely because the file contains sound.
An <Audio N> definition primarily states the audio's role and does not have to name the <Video N> it comes from. State the shared source only when needed to remove provenance ambiguity, for example:
<Video 1> is the source video for the target video edit. <Audio 2> is the synchronized audio track of <Video 1> and is reused in the target video.
like i said in a previous comment, as long as you reference what is what, you don't necessarily need to worry about linked numbers.
@mrweaz That doesn't make it clearer for me. In the base workflow, with the base node you are the one responsible of linking the video images to a video node and the audio of the video to an audio node. In this situation its very clear which number I have to refer them to. My question still stand. With this MiniMax H3 Director node, which of the audio land at #1 if I use a plain audio file + a video that has audio.
@maway it doesn't matter, you just need to tell the model what the audio is. but in your example of having a dedicated audio track and a video with an audio clip, it would be <Audio 2>. but really you should just be referencing or telling the model what audio clip is what.
so lets say i dont know what <Audio> tag to use if i have multiple sources, i would tell the model <Audio 1> is the song <Audio 2> is the dialogue from <Video 1> etc.
The era of wan2.2 is over
Hi, where can i find the MiniMaxH3SigmaShift node, is it also at KJnodes? as i can't seem to find it thanks
Git clone latest Kijai Nodes from github, he did not update the version for now, so it will not be detected from manager
@Darksidewalker is downloading the GitHub as zip and copying it over the old kj nodes afterwards doing the pip install for the requirements.txt incorrect?
properly cloned it this time using cmd, though still missing MiniMaxH3SigmaShift
if you already installed the requirements just copy over
@Darksidewalker thanks for the trouble, still i have no MiniMaxH3SigmaShift node, only got the other minimaxh3sageattn from the install, maybe just out of luck
Updated comfyui as a whole and suddenly it appeared
@alopgamers how do you guys even do that? git pull doesnt solve anything
@LuckyCharmEr from the comfyui manager i just hit update all, after the restart it worked
Did you find a proper solution to this?
With updated nodes im stilling getting the error of the missing MiniMaxH3SigmaShift
Благодарю за workflow. С одной стороны модель интересная, с другой стороны более цифровая, словно симуляция. Ну анимирует она не плохо и звук есть ( из коробки ) 📦
I am unable to install the node MiniMaxH3MemoryEfficientSageAttentionPatch :( please help. I have HP victus 15, rtx4060 8GB
You either have to update KJNodes from the GitHub manually, or you don't have Sage-attention installed.
Delete Kijai Nodes folder, update comfyui to the latest using .bat, git clone Kijai Nodes from its github, restart comfyui, it should download all the missing things. It's the only thing that worked for me.
@mrweaz Thanks, Sage attention is installed. Worked for LTX
@LuckyCharmEr Thanks a lot. That fixed :)
Anyone figure it out how to do proper lip syncing? I want the audio to be the source not a reference. Thanks
Question about the ref2vid function: when adding reference images in the director node, there is the ability to add prompts per image. Would you describe what you want the system to do with that image using that interface or do you still need to reference them using <Picture 1>: syntax in the global prompt?
I just used the global and it was shockingly accurate:
Am I missing something or we can't choose randomise|fixed seed in this workflow, no? I want to use a fixed seed, so how do I do this?
comfyui bug, it does not populate to subgraph atm
you could try and connect a seed generator node to the subgraph and select fixed on generate.
Im sitting here whole day trying to make it work, but im getting "sageattention is not new enough version or could not determine CUDA architecture, cannot apply MiniMax H3 Memory Efficient Sage Attention Patch." error, although i installed last sage attention, i have 130 cuda and matched pytorch, can someone explain what am i doing wrong?
Problem is only with sageattention, everything else works fine, im losing my mind
You possibly using sage-attn 1 not 2
@Darksidewalker im using this one https://github.com/woct0rdho/SageAttention/releases/tag/v2.2.0-windows.post6
Something is not aligned. You can try asking AI to help. Or, do as I did: install a clean ComfyUI environment just for MiniMax. Easy install (not super fast, but easy) - GitHub - Tavris1/ComfyUI-Easy-Install: Portable ComfyUI installer for Windows, macOS and Linux with EZi Desktop app 🔹 Nvidia GPU support 🔹 Pixaroma Community Edition · GitHub
@ArchieMaser yea im starting to think about fresh install, but its kinda personal already xd
I deleted the memory efficient patch node and the flow has worked fine without out. Updated kjnodes and tried to re-add it with no luck.
I won, if anyone needs, you need to do a proper triton installation, and then put two folders from here https://github.com/woct0rdho/triton-windows#8-special-notes-for-comfyui-with-embeded-python in your python_embeded folder
Yeah getting the same error, tried everything including reinstalling Triton and Sage 2.2, nothing works for this node. Every other workflow using Sage 2.2 works fine, so idk what's up with it, a shame
@biggchungus1337 doesnt seem to be sage directly, just the memory patch. Bypass or remove the patch and sage just fine (At least on mine it did).
@srfuentes99 Ah i see, yeah that worked, and sage seemed to shorten the gen time as it should, thanks!
Any idea why the If/Else nodes return these errors?
2 Validation failed
The workflow couldn't validate a connected node.
If/Else Switch - audio
If/Else Switch couldn't validate audio: tuple index out of range
If/Else Switch - images
For some reason some images come out squished, and some are just fine even tho I change nothing in settings, it's set to custom (I2V) and 0.65 balanced. Can someone explain to me like I am 5 what should I tweak in my settings?
And if I choose an aspect ratio which is identical to my image, e.g. square 1*1, it starts generating pixelated and distorted output...
Select Custom?
@skipwestcott as I said, it's set to custom (I2V) and 0.65 balanced. And the images that are not 16:9 (as I've tested) come out squished.
@skipwestcott and when I set a certain aspect ratio, they are no longer squished but now all blurry and distorted.
custom is not exposed, you would need to set this in backend
Just when LTX2.3 was ramping up, Minimax kicked it down a flight of stairs....
Awesome initial iteration of the workflow as always. Definitely painful on a 3060 but with I snuck in RIFE interpol to generate at 12 and was able to get 10 second generations in around 12-15 minutes which is great for a low end pc
Thank you!
Ltx 2.3 was destroyed within minutes.literally everything "except speed" is better.and ostris working on a 4step lora.... so wait few hours or days ;)
Thanks to @Darksidewalker for creating a workflow.there will be more thinks to come from the community.i dont think anybody will use or train ltx.
@hartweizen whats crazy is with how strong the ref to vid is, you don't even need to train loras anymore.... its... an insane paradigm shift
LTX feels like the first version of Will Smith eating spaghetti compared to this. @Darksidewalker great work as always, man. Having an RTX3070 TI with gen times of 5m with strong prompt adherence and no need of loras is insane.
@nekulover24388 I have never been more motivated to scrounge together $1000 to buy a 5070ti
Can someone help me out? I've installed everything as instructed so far and when I run the prompt, it says:
Cannot read properties of undefined (reading 'output')
is there anything I might have messed up that I don't know?
Unfortunate I have no clue what is wrong with this.
Make sure your comfyui and all nodepacks are really up2date.
NO SCALE toggle (on) - throws an error.
# ComfyUI Error Report ## Error Details - Node ID: 1512:2695 - Node Type: MiniMaxH3DirectorGuide - Exception Type: RuntimeError - Exception Message: RuntimeError: Expected 4D or 5D (batch mode) tensor with possibly 0 batch size and other non-zero dimensions for input, but got: [1, 3, 1, 0, 16]
You have to set it off. Since there is no input for the calculator atm.
This is officially the first time since I started using AI that I can say generating a video is fun, fast and with good outputs. I love minimaxH3 already and your workflow is really good.
Only the prompting is a bit complex but I made a workflow on my page to generate it with a simple natural prompt, feel free to steal it and include it in your workflow Dasiwa!
Thank you and have fun! I'm happy if this makes someone happy!
After testing it excessively for several hours, I encountered quite a few outputs with artefacts. No matter what resolution or input image I use, even 3sec 0.83 videos have some baaad artefacts around faces, I compare it to wan and ltx and they have never given me such results. I feel like it's especially visible if you try to use some real images, the deformed eyes betray it all. Any idea how to fix it? I can't believe nobody else is experiencing this...
Not alone
Looking at my gallery of generated videos, I kinda feel like the first frame of every minimaxh3 output is already deformed, as if it's been processed thro something and downgraded in quality. The first frames of all of my ltx, wan videos are crisp and clear (and I am using your workflows for them btw). Could it be an issue with this specific workflow nodes?
so its probably a debateable topic right now as its so early days, but after experimenting since it launched i found that EasyCache dramatically decreases quality (for me at least, guessing its hardware specific) the only actual speedup method that works for myself that gives a good quality output is Spectrum. But Darkside did tests himself and found the opposite. so I'm guessing there's more factors at play. but i would give Spectrum a shot instead of EasyCache and see if it improves for you. again, early days so there's no definitive answer.
EDIT: Others on the Banodoco Discord have said the same thing.
@mrweaz I disabled EasyCache early on cause it produced like major artefacts, but even with sageattention it messes up a lot of things, esp as I found if the initial image is very crisp like a real photo.
@LuckyCharmEr hmm i know that using IMG2VID for this model can produce artefacts on distant shots anyway, it handles close ups waay better. but anyway, i'm using this workflow but i swapped out the Diffusion Loader to Bob's INT8 Loader, EasyCache to Spectrum and my results are decent.
Recommended settings for 16gb like a 5070 ti and 64gb ram? I get oom with sd preset, had to go down to small. 25 steps takes about 20 minutes pre-upscaling with small preset. I am considering going down to 20 steps. Any advice would be appreciated.
i have the same specs as u with sd preset takes 3 mins.
firs great wf, its hard to customize but i found a wy to add auto prompt with images using generate text.
i'm sorry if this was asked before or is in guide but i been playing with the settings and i cannot find a way to make the images to fit or crop if using First frame the video is out like streched to not in correct aspect ratio in original WF there is a Scale Image to Total Pixels that fixes that but i cannot find the right combination here. thanks in advance
edit i "fixed" by adding the nodes from original workflow and connecting the first image to them., then connecting width and height node to the director.
If you don't want to do that on every update, a simpler way is to select CUSTOM aspect and in the backend you can link the Custom_aspect_width and Custom_aspect_Height from th DaSiWa Resolution Scale Calculator node to the frontend so you can easily change the resolution without going in the backend each time.
Absolutely amazing workflow - right out of the box! Major props @Darksidewalker - thank you! Holy crap Minimax H3 is amazing... Runs great on my 5090!
Glad to hear! Do some nice art! 😸
Solid work. One bug I noticed: On the MiniMax H3 Director node when using the REF2VA mode with one video and multiple pictures, everytimes you click on a picture it moves it around, sometime swapping it around with the other picture. That makes Picture 1 becomes Picture 2 and the references in the caption no longer match. Have to constantly juggle around with it to put them back in their right place.
Thank you for reporting the bug!
Fixed in NodePack v0.4.5 ✌️
Great workflow, the director makes it really simple to add and pre-tag references.
If I could beg a feature request for the next version: A clear button that only clears the references and not the prompt would be really useful.
In the meantime the note node is your friend ;)
the DaSiWa Director loads audio through soundfile/libsndfile, which — unlike the core ComfyUI LoadAudio node (which uses PyAV/FFmpeg) — doesn't support .m4a/AAC at all.
why use soundfile/libsndfile if doesn't support .m4a?
Thank you for reporting the bug!
Fixed in NodePack v0.4.5 ✌️
If you're on the fence about this workflow or MiniMax-H3 in general, just do it. If you have trouble with sageattention or anything like that, bite the (very soft and easy) bullet and do a fresh install with DaSiWa's comfyui installer, and just copy your models folder over from your current install. MiniMax-H3 really is absolutely insane right out of the box, and finetunes and loras will only make it even better in the coming months or even potentially years to come. This is the brightest gooning timeline, lads.
agreed, I finally fixed my sageattention that i've been putting off for months for this workflow :P
there is a bug where when you drag to reorder the pictures, the title stays in order but the reference is wrong. Swapping Picture 2 and Picture 3, the UI will show the correct order but the reference carries overs so the prompt <Picture 3> is actually Picture 2
Thank you for reporting the bug!
Fixed in NodePack v0.4.5 ✌️
As far as I can tell, there is no way to specify that you want to use L2VA mode when using one image. It always treats it as I2VA. The workflow is great otherwise, but not being able to use an image as an endpoint without a starting image to pair with it really limits what you can do. Am I missing something?
It can do last frame and zero start, update my nodepack.
Set the image slot 2, it will use it as L2VA.
you can use the starting prompt: How the reference pictures align with the target video — <Picture 1> (from [Shot 1]) aligns with the 5.00-second mark of the target video. or whatever lenght you are using
I get an error when I try to activate the DASIWA RTX upscaler. (all nodes are installed and everything is up to date.)
RuntimeError: NVIDIA RTX VFX (nvvfx) module not found. Please ensure NVIDIA RTX Video SDK / Broadcast SDK is installed and the 'nvvfx' package is in your python path.
