H3 Throat Time — MiniMax H3 LoRA
H3 Throat Time is a specialized audiovisual LoRA trained on a highly curated data set of oral sex video clips over a 700-step, 50-epoch training process, designed to enhance throat-focused adult scenes, motion, and visual realism.
Recommended Strength: 1.0
Included Versions
I'm releasing two checkpoints:
Epoch 36: The most versatile and balanced version overall. Recommended as the starting point.
Epoch 48: Slightly more visceral and intense in its output. The differences are subtle and largely stylistic rather than dramatic.
Trigger Word & Prompting
Trigger word: thr0at
This LoRA responds particularly well to detailed, descriptive prompting. More comprehensive explanations of the intended action, movement, pacing, and visual characteristics generally produce better results than short, generic prompts.
For best results, use the trigger word alongside clear, natural-language descriptions of what you want the model to accomplish. The more precisely you communicate the intended scene and motion, the better the LoRA tends to interpret your instructions.
So far testing has confirmed it handles the following scenarios quite well.
Solo Wet Slobbery Blowjobs
Assisted Wet Slobbery Blowjobs (character guides another character to blow a character)
Alternating 2 person blowjob (prompt language must be strict to avoid them both trying to blow the person at the same time)
Fast Paced Visceral FaceFuck (with or without cum and slobber)
Coughing while deepthroating penis in mouth
Coughing while coming off penis (with or without semen)
Note on Semen: Model does better when semen is in source picture. H3 tends to default to pancake batter consistency without a visual cue. You can also prompt it away by asking for semi transparent seminal fluid or cloudy transparent semen.
Audio — A Major Motivation Behind This LoRA
One of my primary motivations for creating H3 Throat Time was to address the shortcomings of MiniMax H3's native audio generation, particularly its tendency to produce unnatural, crunchy Foley effects and unconvincing wet mouth sounds.
I wanted to push beyond those characteristics and improve the overall audiovisual experience rather than simply training another motion-focused LoRA.
Standard 20-step generation (recommended):
The LoRA performs best without Turbo acceleration, using a conventional 20-step sampling process. This provides the strongest overall balance of animation fidelity, audio characteristics, and the stylistic qualities learned during training.
Turbo generation (6–8 steps):
Turbo is supported as an alternative for faster generation. In my testing, it generally preserves most of the animation fidelity, making it a practical option when rendering speed is a priority.
However, expect some regression toward H3's baseline audio characteristics when using Turbo. The original crunchy Foley effects and less convincing wet mouth sounds may become more apparent, as Turbo can diminish some of the audio improvements learned by the LoRA.
For Turbo workflows, I recommend 6–8 sampling steps.
Bottom line: If audio quality is important to your generation, I strongly recommend running the LoRA without Turbo at the full 20 steps.
Anatomy & Reference Guidance
The training dataset included footage featuring male anatomy. When male anatomy is fully visible and properly referenced in the source shot, the LoRA can generally handle it without additional anatomy assistance.
However, despite being trained on this material, I do not recommend relying on this LoRA alone to generate male anatomy from scratch. For best results, use either a complete visual reference or a dedicated anatomy-trained LoRA.
Compatibility & Testing
Currently validated:
I2V (Image-to-Video)
FL2VA
Not yet validated:
T2V (Text-to-Video)
REF2VA
T2V and REF2VA compatibility remain unknown. I make no guarantees regarding performance or functionality in those modes at this time.
Recommended Starting Point
Start with Epoch 36 for versatility, then experiment with Epoch 48 if you prefer a slightly more visceral aesthetic. For the most faithful reproduction of the LoRA's learned motion and audio characteristics, use detailed prompting and the standard 20-step process.
Credits
Credit to @fatberg_slim for several really great Wan videos that were used in the training set for this lora. The remaining data set was from sourced via Pornhub - LOL
Description
FAQ
Comments (7)
Appreciate the lora. But it's too bad about its issue. Many if not all H3 blowjob loras suffer from making it look like the woman is trying to EAT his penis. This kind of chewing motion.
It looks unnatural and a bit scary lol
this. plus the monster lower lip. Looks like someone needed to dump surplus of silicone somewhere.
It's more of a shortcoming of the base model. My perspective is it's an improvement over the baseline functionality. As a hobbyist, that's good enough for me. if people have some fun with it that's even more bonus. I was generally pleased with the end result. Thanks for providing feedback.
@lucilious201 you're welcome to curate your own data set and give it a shot. It's not at all simple. Thanks!
this lora need more learning steps. I have trained dildo throat lora (cant upload cause low resolution video = bad output) but they stops eat penis/dildo after 2000 steps. (depend of you your data set size ofc)
Sounds and motion with this lora is great. The big issue it has is the lips around the shaft aren't that realistic. But I've found if you add the other facefuck lora at 0.2 - 0.3 that does enough to fill in the gaps while keeping the audio and other strengths of this lora.
Thank you. I wouldn't be against a longer training run at a higher resolution. I'm generally curious about training it at .5MP instead of the standard recommended .25. I know that will increase the VRAM and time with the set but it might correct some of the lip stuff. I will of course investigate. Thank you!