Lazy MiniMax H3 — All-in-One (Auto-Switching)
Pick the workflow type. Everything else follows.
This is a Lazy MiniMax H3 graph: one selector, one conditioner path, one sampler stack. Set T2V / I2V / FL2V / R2V and the all-in-one nodes route models, frames, refs, and audio for you — no muted branches, no duplicate samplers, no “which UNET did I leave on?”
The Lazy idea
Lazy here means the workflow does the bookkeeping.
You choose the workflow type once.
Lazy MiniMax All-in-One builds the right conditioning for that mode.
Lazy Model Switcher picks fl2va (text / image / first–last) or ref2va (reference).
Lazy Image Loaders only emit what that mode needs; unused roles hard-gate off so optional sockets stay empty.
LazyPrompt can expand a short idea into a MiniMax-ready prompt — then get out of the way.
Everything is modular. Nodes are grouped for clarity, not locked layout. Drag sections wherever you like; the wiring stays intentional.
Modes at a glance
ModeWhat you loadDiffusion
T2V
Text only
fl2va
I2V
First frame
fl2va
FL2V
First + last frame
fl2va
R2V
Reference images (+ optional audio)
ref2va
Change the selector → loaders, MiniMax, and UNET routing update together.
Prompt enhancement (optional, Lazy-friendly)
You can run raw text straight through (Prompt Engineer bypass) or enhance first.
Recommended: LM Studio (local LLM)
Workflow can load the model, enhance, then unload when finished.
Typical enhance time: about 5–10 seconds.
Best quality/speed balance for day-to-day Lazy runs.
Vision-capable LM Studio models work well with I2V first-frame grounding.
Alternative: TextGenerate (CLIP)
Uses a wired Comfy CLIP / Qwen LLM — no second app.
Works, but slow — often on the order of ~1 minute vs a few seconds with LM Studio.
Fine if you refuse a separate LLM server; otherwise prefer LM Studio.
MiniMax-oriented skills (I2V / FL2V / R2V) are included so enhanced prompts speak H3’s language (<Picture N>, first/last alignment, reference jobs, soundscape fields).
What’s in the box
Lazy Global Selector — single mode control
Lazy Image Loaders — role + hard-gate by mode
Lazy Subject & Scene Automation — optional subjects/scenarios/LoRAs
LazyPrompt — Prompt Engineer — LM Studio / HF / TextGenerate CLIP
Lazy MiniMax All-in-One — T2V ↔ R2V conditioner auto-switch
Lazy Model Switcher — fl2va ↔ ref2va
Native MiniMax H3 decode → video + stereo audio → save
Built for people who want one graph, not four copies of the same sampler.
Requirements (short)
ComfyUI 0.30.0+ (native MiniMax H3)
Vsaan212-workflow-utilities (Lazy nodes)
KJNodes (Set/Get, helpers), plus usual Comfy core loaders/sampler/save
Models from Comfy-Org/MiniMax-H3: fl2va + ref2va UNETs, Qwen3-VL encoder, video VAE, audio VAE
Quick start
Install custom nodes + download the H3 files into
diffusion_models/text_encoders/vae.Open the workflow; set Lazy Global Selector to your mode.
Drop images on the matching loaders only.
Write a short idea — enhance with LM Studio (recommended) or bypass for raw text.
Set duration on MiniMax All-in-One → Queue.
Description
Fixed some oversights on wiring.
Workflow now only uses Vsaan212 and KjNodes
if you downloaded before, update vsaan212 nodes to get the new nodes.
added lazy documentation nodes. use the drop down to select this workflows documentation.
3 sample workflows for I2v, First last and reference to show how you can rearrange the nodes for different workflows.
Added hugging face link to the "Clip For prompt enhance" node
https://github.com/vsaan212/Vsaan212-workflow-utilities/ documentation is more in depth.
any suggestions? let me know
Feel free to use any of this in your nodes or workflows.


