CivArchive
    H3 Eros Max - beta2
    NSFW

    beta 5:

    Use TURBO-hybrid_int8 by default. Other options are for experimenting. Full versions here.

    beta_5 is built with a different normalization technique. It's also built off 7 different concept-grouped grafts and consensus merges from over 20 Loras, no full direct lora merges. It allows for a non-turbo version which is available. The files with TURBO have a hybrid turbo-delta fusion baked in saving 4.2 gigabytes of memory instead of having to load both ref/fl turbos. Other concept Loras will also load on top more readily, usually needing lower strength 0.2-0.6. For t2v and even some i2v you should 100% be loading mystic_v4, anatomy enhancer, or a related concept lora on top.

    Also consider using the non-turbo with PDD, 6 warmup and 4 PDD turbo steps.

    Prompt is the entire key to quality and success. Any issues you'd want to blame on a model or workflow, you can go ahead and take a look at the prompt instead. Describe things cleanly and literally:

    ❌ He puts his penis in her pussy, make hot sex

    ✅ Live-action pornographic explicit sexual and sensual intimate POV recording: The man slowly moves his lower body forward with his finger on the base of his penis shaft. The tip of the penis slowly disappears into her pussy hole and her wet labia open allowing it entry. He keeps swinging his pelvis forward until his crotch touches hers. Then he backstrokes and starts a repeated thrusting sex motion. The sex act makes her whole body recoil into the couch, bouncing her breasts wildly. She stares lustfully into the camera.

    Prompt in temporal sequence on a linear time flow. The longer and more embellished the prompt, the better the output will be usually. 100% use some kind of LLM enhancer; Grok is the one I use since he has a skill preset for reference prompting.

    The version just labeled 'hybrid' is non-turbo. Non-turbo full step audio will always be better than audio on turbo versions.

    Some working sampling setups:

    • er_sde/beta57 4-6 steps (seems to be best preservation of style and reduced drift)

    • res_multistep/simple 6-9 steps (good motion quality, beta will be the sharpest fast motion at 8-9 steps but will give a plastic/burned look)

    • LCM/simple 6-8 steps (best turbo audio)

    • Euler/simple 4-8 steps

    Next version should be mainly powered by the first round of Sulphur training. From what I've seen it will fix most missing audio and missing concept issues.

    Updated - Full credits to these lora makers for having deltas involved in a consensus merge in beta_5 version:

    alcaitiff, MisticRain69, diogod, FourBunny, HearmemanAI, tazmannner, simonishere, QualityControl, blo01, ComfyTinker, kermitfrog1202, qdr1en, coachbate, salttaro, alternative_penguin, misterxrex, freek22, definitelynotadog


    beta 4:

    rebuilt on beta3 config with some loras changed out for newer versions. Turbo was consensus merged to construct a ref/t2va hybrid turbo lora. That's the key element to the merge, there is no non-turbo version. The non-turbo version is bad and doesn't work since the turbo weight is normalizing the merge by it's strength. That custom turbo merge will need further improvement. This one preforms video, reference, and motion well in 6-8 steps without the issues from beta3, but audio needs shift configuration.

    Use sampling like Euler/simple 8 steps with sampling shift - 12 video/ 7+ audio. LCM/simple or beta with 6-8 steps and no shift can also be better for audio and drawn styles.

    Audio is lackluster and it's becoming somewhat apparent that H3's integrated audio is not good, like terrible actually. Any future multimodal models should avoid integrated audio if they intend to open source. I have almost no control over how the audio works inside the model. Don't post about it. I focused on motion and prompt response and of course when I get those working well, the audio ends up bad, go figure. I'll look at what kind of different turbo configs can enable more audio crispness to come back or likely will have to wait for Sulphur to replace MysticXXX which is contributing to the audio quality drop.

    beta 3:

    Rebuilt on the delta1024 reference hybrid model. Does t2va and reference. Treat i2v as a single image reference, don't do i2v prompting. Silver's merged turbo is integrated and tuned for 6 steps with simple schedule with the extra _emb layers tacked on to the model, not sure if they're needed.

    Samplers like er_sde or multires or other turbo sampling setups work. Best imo is just er_sde/simple 6 steps, no shift, no spectrum, no cache, only a comfy_kitchen backend selection node. 3 video references in a 15+ second outputs can be done in under 10-15 minutes now on larger cards with no extra cache or quality hit needed.

    Many (like a lot, all the good ones) on-site loras were fully combined into a consensus weighted merge with ranked drop-out to form an initial part. That merge is put against new, more powerful Wan and LTX grafts as a blend/reshape that uses the loras to consensus shape the grafts, but it also allowed some of the better loras clean pass-through. This is not a linear list of loras just merged. The main element that can present most is probably MysticXXX which was given the most pass-through weight since it's just good--and all 3 release steps of that are inside it. However, they're all-combined with a ton of other loras with agreement and consensus of shape and then it's only reshaping the grafts parts. The results of that are the actual weighted loras that are used to make the model.

    Due to drop-out and consensus merge, pretty much all of the loras can all still be used easily on top if needed, and might work better even. All of this was only done to create a large rank dummy lora similar to what sulphur data will look like as a lora or extracted lora so I can start looking at how to apply it cleanly.

    It's definitely not a few on-site loras that are linear merged, uncredited, and then renamed with some emojis. I'll only do this until sulphur tuning steps are in my hands and I can work with more targeted and shifted stuff, plus I was tired of waiting and I wanted fast easy i2v.

    There is one quirk of the hybrid h3 usage: don't use it i2v. It should either be always used in reference prompting mode, or t2va prompting mode. Even if there is just one single image input it needs to be used as reference and prompted in the ref2va format. If you run an underdeveloped or manually written prompt you will get odd outputs, random camera changes, and blue lighting color shifts when you use the i2v prompt style.

    Full credits to these lora makers for being involved somewhat in beta3 version:

    alcaitiff, MisticRain69, diogod, FourBunny, HearmemanAI, tazmannner, simonishere, QualityControl, blo01, ComfyTinker, kermitfrog1202


    Beta2 and previous:

    This started a finetune-by-graft. Or maybe a GST - grafted shift of transformer (cross-architecture). I made both up, because there aren't any projects that have done it that I know, except one reddit post that made me look into it. I experimented with Wan and LTX on the side which led to the initial LTX Eros scripts that became what powered this, all before H3 ever came out. It seems like unified unbiased models like MMH3 can technically take attention influence from any other DiT without breaking if done correctly. Anima, Krea2, LTX, Wan2.2, Flux1 were all tried out, configs tested, about ~40 hours maybe of working in the dark without any paper or technical documents from Minimax. Eventually I developed linear-magnitude blend application and specific block and head gate targets allowing for a smoother graft on an attn-triplet-unfused version of H3 output as a patch file. That sent to lora extraction, then merged to checkpoint at taste. This is a merge but a merge of LoRas I extracted that interact to produce this current shift. I saved 5 ponds of water by recycling data in a few minutes on a single card instead of toasting a server up.

    Turbo not recommended yet for i2v, especially when used with other LoRas. T2V use with turbo is better. Use 20-25 steps normal sampling with no dialogue, 25 steps with dialogue along with cache nodes and attn modes. More steps over 25 are not neccessarily better, and can be worse. Use full int8: int8 model, int8 VAE (if it doesn't crash comfy), int8 qwen3vl along with current cache or attn mode nodes. For smaller cards: quants, macOS ports, and Wan2gp support will likely appear on huggingface but not from me.

    Known quirks:

    • Audio difference v.s. Base - This model's audio changes come from attention shifts seeking alternate audio pairing. Attn triplets were unfused before graft, both standard and triplet q_attn was grafted holding about maybe 10-15% audio influence, attn_k was frozen and MLP fc2 layers were untouched resulting in minimal audio interference. This was the main issue with the entire transformer graft and protecting audio. However this version is slightly louder overall than the base model.

    • Low resolution detail smearing - Some finger digits and fine motion will smear more at low resolution, also a problem in base model. As memory use gets more efficient increase resolution or work on the composition to get around it.

    • Odd outputs - This can attempt certain concepts more liberally than base model, but that can lead to some undesirable outputs in bad prompting and certain contexts. Data shift comes from completely different transformers and architecture. This shouldn't even work, so it is what it is.

    This model is not dedicated to NSFW as that would violate community license agreement. Sure it can do it, just like base. Any NSFW generations are purely the result of advanced reasoning and tokenization resulting from experimental changes. All terms from the H3 community license also still apply to the users of this version. Don't be a dumbass.

    H3 usage still requires very intense prompting for maximum effect. Every motion, every interaction, every sound plainly and fully described. Not with slang terms; with proper actionable words that can be tokenized. Refer to the h3 developer prompting guide, hand that .md file to an LLM or Chat agent and have them enhance or refine prompts along the released H3 developer prompt guide styles using the model's tag system. Certain concepts can be made from pure token reasoning. Consult the prompts in my previews to see certain physical descriptions that I use for some things. When using enhancement give the agent feedback about any issues in the generation and get them to describe motions in alternate fashion, or manually edit it yourself adding a negative like "no X, no Y". Still requires prompt refinement and trial/error for best outcomes.

    Sulphur Project has 10k banked to attempt actual tuning. Right now training pipelines are sub-optimal. As always Eros is my personal side project, and this beta was also essentially a speed-run of finetuning, figuring out exactly in what configurations and target areas do you get helpful/harmful changes in the model. This is also a proof-of-concept of what and where to target while leaving the reinforcement quality of base unharmed by being additive.

    https://huggingface.co/TenStrip/10Eros-Max

    https://ko-fi.com/tenstrip

    Description

    Initial release-worthy grafted shift tune.

    FAQ

    Comments (47)

    283782122383Aug 16, 2026· 1 reaction
    CivitAI

    great!

    TTTTT55Aug 16, 2026· 1 reaction
    CivitAI

    Aahhh yeee!

    KiraNuggetAug 16, 2026
    CivitAI

    Looks amazing! Cant wait to try.

    tenstrip
    Author
    Aug 16, 2026· 2 reactions
    CivitAI

    This flv2a can still be used in reference workflow but only with image inputs. The ref model is still needed for actual full video and audio reference.

    Good info! The reference mode is awesome, but the reference model needs more work.

    trankhacvinh1991Aug 16, 2026
    CivitAI

    Wow, can't wait to try it. Amazing work

    mycoreAug 16, 2026· 2 reactions
    CivitAI

    wow h3eros version nice, h3max look good but so heavy i think, and some test say take more time to generate than ltx.. i think i will still use the ltx

    tenstrip
    Author
    Aug 16, 2026· 5 reactions

    My better LTX Eros gens with the extra sampling steps took about 1-5 minutes for really long ones. This model takes 1-5 minutes as well. All you have to do is keep resolution and length relatively low, there is a point where you start offloading too much because the latent is too big and it slows down.

    gambikules858Aug 16, 2026· 5 reactions
    CivitAI

    plz ref2v

    Denkaichi_XAug 16, 2026
    CivitAI

    ypikayey

    alphaatlas100844Aug 16, 2026· 5 reactions
    CivitAI

    Can you make an A/B test or two vs minimax H3? Same prompt, same seed?

    I find this fascinating but... I'd like a more objective illustration of what the graft actually does, if that makes sense.

    tenstrip
    Author
    Aug 16, 2026· 4 reactions

    I did probably 20 A/B scenarios that's how I even know I had a good configuration going. The differences aren't profound because it's the same model but they diverge more and more when you go in the direction of the donor model. It's like bread v.s. bread with some butter on it.

    Kingp0dd529tAug 16, 2026

    Can we see those

    tenstrip
    Author
    Aug 16, 2026

    @Kingp0dd529t here's one but it's NSFW. The differences aren't substantial but they matter, and it's small details and motion fixes across pretty much everything. Until I find ways to keep pushing it further. The goal isn't to do much of anything to the base model besides fix it up. It's pretty good already and there's loras for anything else.

    alphaatlas100844Aug 17, 2026

    Appreciate the upload, that's interesting.

    It's... very minimally altered vs the base model, at least in those examples. Which is good, I suppose.

    tenstrip
    Author
    Aug 17, 2026

    @alphaatlas100844 Yes it is good. I've tried stronger weighting and it starts doing a bit too much. It's already maybe a bit too much Wan influence, if you notice the slower motions (Wan is 16 fps)

    Light7799Aug 17, 2026

    @tenstrip Have you noticed if this one is better at any concepts the original one couldn't do as well? I kind of understand a good bit of ai and training and blocks, seeds, etc... but I've never heard much about grafting. Very curious how to even go about grafting and if it can introduce or reinforce lost knowledge or if it's more of a detail fixer and/or follows prompts better.

    tenstrip
    Author
    Aug 17, 2026

    @Light7799 It's something I guess I started, so yeah you haven't heard of it. Not sure what it is besides a training-less way to change the model's self attn. It can work on any model tbh, I tried H3 -> LTX 2.3 as well. Works way better when it's actually shared architecture of course, the last two versions of Eros were made like that.

    scarlettdays21Aug 16, 2026
    CivitAI

    Hi I already try your ltx 2.3 eros it work great on my system so will this also work with 64gb ram and rtx 5070ti ? it almost 40gb in size so i would assume it will not...

    tenstrip
    Author
    Aug 16, 2026· 1 reaction

    The int8 model variant is under it that's the one to use, and smaller than ltx2.3 actually but needs 25 sampling steps.

    scarlettdays21Aug 16, 2026

    @tenstrip Thanks i will try it now

    gigcamellos2023710Aug 16, 2026· 1 reaction
    CivitAI

    any recomendations for a good h3 workfkiw ref to vid

    tenstrip
    Author
    Aug 16, 2026

    I've just started using the reference model today. I was focused on i2v and t2v so far. Both of the default comfy template workflows are pretty good for the model they just need the speed-up nodes put on them.

    KiraNuggetAug 17, 2026

    Dasiwa's fore sure

    suekapoAug 16, 2026
    CivitAI

    Very good model, my only problem is that each step takes much longer, 1st step starts with 180it/s, at 4th step at 300it/s, I can't figure out why, other models do 112it/s with the same setting, any idea why this is happening, am I missing something? Thanks!

    tenstrip
    Author
    Aug 16, 2026

    High it/s like that isn't actually good you're bottlenecked at offloading using too much memory. The int8 on this one is the first one I got and it's 1.1g larger than the base pruned int8. I'd add the normal full row wise one but I don't have one, but you need to turn size slightly or turn down time by a second.

    suekapoAug 17, 2026

    @tenstrip Yea I know I'm offloading, but I'll take it for the better resolution, also works good with Turbo LoRA, I lowered the resolution and don't have that issue anymore with your model. Didn't think this 1gb would cause that, thanks! Testing the model now

    dav79mail156Aug 17, 2026

    @suekapo What resolution do you use and what is the turbo lora?

    KiraNuggetAug 17, 2026
    CivitAI

    Some one just told me to use FL2V model even when doing ref2v and it works great. No lora incompatibility, and the quality and speed of fl2v model!

    chaiwala008215Aug 17, 2026

    what prompt u use to make the jump / cutscene / ?

    KiraNuggetAug 17, 2026· 2 reactions

    @chaiwala008215 at 0:05.000 the camera cuts to [Shot 2]

    tenstrip
    Author
    Aug 17, 2026

    I have to remake both grafted pieces from scratch using the ref model, so that's in-progress.

    CyclopsGERAug 17, 2026
    CivitAI

    Is Gay stuff possible? Its often ends up in "hetero" sex.

    tenstrip
    Author
    Aug 17, 2026· 1 reaction

    I feel like the model follows prompts so literally that if you describe it enough with the right i2v start, then yeah it could work.

    CyclopsGERAug 17, 2026

    @tenstrip I am using text to video.

    tenstrip
    Author
    Aug 18, 2026

    @CyclopsGER Yeah you'll need gay loras and a penis lora then.

    amazingbeautyAug 17, 2026· 3 reactions
    CivitAI

    fp8

    why458Aug 17, 2026
    CivitAI

    Love your LTX model and excited to play and learn more about this one! Is there a place where I could look up some of the things this model was trained to do to streamline the learning and exploration process?

    tenstrip
    Author
    Aug 17, 2026

    There was no training. This is altering the model's behavior more deeply underneath it's reinforcement with attention grafting on block heads. That's part of the goal/experiment since with a model this good the last thing you'd want to do is just SFT it on a couple thousand clips and negatively impact all seeds. I don't think a model like this without a proper training pipeline or an alternative version is trainable like older non-deterministic diffusion models were.

    FantasticFeverAug 18, 2026
    CivitAI

    int 8 seems to run a lot slower than H3 int8 convrot pruned

    tenstrip
    Author
    Aug 18, 2026· 1 reaction

    Yeah I'm trying to get a normal int8, the skip edges runs mixed or something slowing it down.

    nailiang002188Aug 20, 2026
    CivitAI

    I saw u upload ref2va version on hugging face. will u upload on civitai? cuz download from HF is so slow for me. it's great work! thanks :D

    Taloco22200Aug 24, 2026

    Use hf_transfer (pip install).
    Ask whatever LLM you use and it'll explain how to set it up. Multi-threaded, way faster than the browser download.

    nailiang002188Aug 25, 2026

    @Taloco22200 thx,bro.I will try.

    westlonger356Aug 23, 2026
    CivitAI

    请问是否有可能做一版LTX2.5的版本。另外LTX2.3的1.5版本似乎是已经在hugging face上发布了么~

    tenstrip
    Author
    Aug 23, 2026

    LTX 2.5 versions I've tried are just worse versions than the 2.3 models. Comfyui still hasn't added the models new features, and it really needs a Sulphur retrain on LTX2.5 to be full quality.

    Checkpoint
    MiniMax H3

    Details

    Downloads
    2,864
    Platform
    CivitAI
    Platform Status
    Available
    Created
    8/16/2026
    Updated
    10/5/2026
    Deleted
    -