Cinematic Style V2 for MiniMax H3
Cinema V2 is a major expansion of my broad cinematic style LoRA for MiniMax H3.
The dataset now contains 1,852 cinematic images with 1,852 matching detailed captions — 3,704 files in total. Version 1 already provided a strong general foundation for realistic film photography, framing, lighting, lens character, depth of field, production design, atmosphere, realistic materials and organic film texture. Version 2 keeps that foundation and adds much more deliberate camera-position and composition training.
## What Is New in V2
The dataset was expanded from 1,201 to 1,852 image-caption pairs, adding 651 new cinematic examples.
New dedicated categories include:
- Aerial shots
- Clean single-subject compositions
- Dutch angles
- Establishing shots
- High-angle shots
- Low-angle shots
- Over-the-shoulder compositions
- Overhead shots
These additions improve the LoRA's understanding of camera height, viewing direction, spatial depth, visual geometry, subject placement, negative space and environmental scale.
The original dataset already covered the main shot sizes:
- Extreme Close-Up
- Close-Up
- Medium Close-Up
- Medium
- Medium Wide
- Wide
- Extreme Wide
Together, both parts now provide a much broader cinematic vocabulary. V2 is better suited for quiet feature-film scenes, atmospheric openers, dialogue scenes, character moments, environmental storytelling, news-style cinematography, trailers, science fiction, period drama, crime, romance, horror and large-scale landscapes.
The captions place strong emphasis on:
- exact framing and subject position
- motivated key, fill, rim and practical lighting
- lens character and depth of field
- realistic skin, fabric, glass, metal and wet-surface response
- foreground layering and environmental depth
- controlled color relationships
- film grain, halation, bloom and highlight rolloff
- atmosphere, weather and production design
This is a cinematography style LoRA, not a character LoRA and not one fixed movie look. The genre, people, clothing, environment and action remain controlled by your prompt.
## Activation Tag
Use the original activation tag exactly as trained:
ASTROCINEMAV01K2T
The K2T wording is intentional and must not be changed. This release is trained for MiniMax H3.
Place the trigger near the beginning of integrated_multimodal_description.
Example:
```text
integrated_multimodal_description: [Shot 1] ASTROCINEMAV01K2T depicts a live-action cinematic scene in a native 16:9 widescreen frame, photographed on a 50 mm anamorphic lens with motivated tungsten practical lighting, cool window fill, realistic skin texture, shallow depth of field, fine 35 mm grain, subtle halation and controlled highlight rolloff. No cuts; one continuous shot throughout. Describe the complete subject, wardrobe, environment, blocking, camera movement and spoken dialogue here.
overall_soundscape: Describe the synchronized ambience, physical sounds and nonverbal character sounds here.
non_diegetic_music: Describe the score and instrumentation here, or use N/A when the scene should rely only on natural sound.
```
## Important
The LoRA strengthens the visual cinematic language: framing, camera placement, perspective, lighting, lens feeling, production design and film texture. MiniMax H3 itself remains responsible for temporal motion, acting, speech, sound effects and music, so those elements should still be described clearly in the prompt.
For the best results, do not write only “cinematic.” Direct the scene like a filmmaker: specify the shot size, camera height, angle, lens, movement, light sources, subject blocking, environment, atmosphere and sound.
_______________________________________________________________________
# Cinematic Style LoRA for MiniMax H3
A broad cinematic style LoRA for MiniMax H3, trained on roughly 1,200 carefully captioned cinematic stills.
The goal is not to force one specific movie look. The dataset was built to teach a wider cinematic language: realistic film photography, composition, framing, perspective, motivated lighting, lens character, depth of field, production design, realistic materials and skin, color relationships, atmosphere and organic film texture.
## Activation Tag
Use the original activation tag exactly as trained:
ASTROCINEMAV01K2TYes, the tag still contains K2T. That is intentional because it is the original learned activation tag from the dataset. This release itself is for MiniMax H3.
Place the tag near the beginning of style_and_tone.
Example:
style_and_tone: ASTROCINEMAV01K2T. Photorealistic neo-noir crime drama, motivated practical lighting, shallow depth of field, restrained color grading, realistic skin texture, subtle halation and organic fine film grain.## How I Recommend Prompting It in MiniMax H3
Do not rely only on the word "cinematic".
Tell H3 what kind of cinematography you actually want:
- time of day
- dominant and secondary light sources
- practical lights visible in the scene
- camera distance and framing
- camera movement
- lens feeling / depth of field
- realistic skin and material response
- environment and production design
- weather and atmosphere
- color relationship
- film grain, halation or highlight rolloff when useful
The LoRA works best as a broad cinematic foundation, while the rest of the prompt defines the specific genre and shot.
## MiniMax H3 Prompt Structure
I recommend separating the prompt into:
- style_and_tone = visual language and cinematography
- continuous_shot or multi_shot_sequence = action, acting, camera and timing
- audio_design = music, sound effects and ambience
- explicit dialogue lines whenever characters speak
H3 is very good at combining visual direction with audio, so do not forget the audio side when the scene needs it.
For example, a crime scene can include rain, passing traffic, a low score and dialogue. A fantasy battle can include orchestral music, weapon impacts, footsteps and creature sounds. An advertisement can use voice-over, product Foley and a clean commercial music bed.
## General Tips
For realistic cinematic results, avoid contradictory style language. If you want realism, do not mix the prompt with vector-art, cel-shading or illustration terminology.
Use physically coherent lighting. Instead of listing random light colors, describe which source is dominant and how secondary light affects the subject.
Keep skin and materials realistic. Describe pores, tonal variation, fabric response, wet surfaces, polished metal, glass or practical reflections only when they are relevant to the shot.
For fast action, keep the movement readable. One clear action per beat usually works better than trying to force too many unrelated events into a few seconds.
For trailers and advertisements, multi_shot_sequence is extremely useful because each cut can have a clear purpose.
For music-led videos, describe the music as the master timeline. Shot and camera cuts may use timecodes, but avoid forcing sung lyric lines into strict timed segments.
## What to Test
The LoRA is useful for a wide range of H3 projects:
- crime and neo-noir
- period drama
- romance
- action
- science fiction
- fantasy RPG cinematics
- realistic game trailers
- commercials
- beauty advertising
- product videos
- music videos
- atmospheric character scenes
- cinematic environments
The attached example videos were generated with MiniMax H3 using this LoRA. I also included a separate prompt pack containing 15 complete H3 Text-to-Video prompts across film, games, romance, action, trailers, advertising and music-video scenarios.
Feel free to modify the examples and push the LoRA into completely different genres. If you create something interesting, I would love to see the result.
Description
Cinematic Style V2 for MiniMax H3
Cinema V2 is a major expansion of my broad cinematic style LoRA for MiniMax H3.
The dataset now contains 1,852 cinematic images with 1,852 matching detailed captions — 3,704 files in total. Version 1 already provided a strong general foundation for realistic film photography, framing, lighting, lens character, depth of field, production design, atmosphere, realistic materials and organic film texture. Version 2 keeps that foundation and adds much more deliberate camera-position and composition training.
## What Is New in V2
The dataset was expanded from 1,201 to 1,852 image-caption pairs, adding 651 new cinematic examples.
New dedicated categories include:
- Aerial shots
- Clean single-subject compositions
- Dutch angles
- Establishing shots
- High-angle shots
- Low-angle shots
- Over-the-shoulder compositions
- Overhead shots
These additions improve the LoRA's understanding of camera height, viewing direction, spatial depth, visual geometry, subject placement, negative space and environmental scale.
The original dataset already covered the main shot sizes:
- Extreme Close-Up
- Close-Up
- Medium Close-Up
- Medium
- Medium Wide
- Wide
- Extreme Wide
Together, both parts now provide a much broader cinematic vocabulary. V2 is better suited for quiet feature-film scenes, atmospheric openers, dialogue scenes, character moments, environmental storytelling, news-style cinematography, trailers, science fiction, period drama, crime, romance, horror and large-scale landscapes.
The captions place strong emphasis on:
- exact framing and subject position
- motivated key, fill, rim and practical lighting
- lens character and depth of field
- realistic skin, fabric, glass, metal and wet-surface response
- foreground layering and environmental depth
- controlled color relationships
- film grain, halation, bloom and highlight rolloff
- atmosphere, weather and production design
This is a cinematography style LoRA, not a character LoRA and not one fixed movie look. The genre, people, clothing, environment and action remain controlled by your prompt.
## Activation Tag
Use the original activation tag exactly as trained:
ASTROCINEMAV01K2T
The K2T wording is intentional and must not be changed. This release is trained for MiniMax H3.
Place the trigger near the beginning of integrated_multimodal_description.
Example:
```text
integrated_multimodal_description: [Shot 1] ASTROCINEMAV01K2T depicts a live-action cinematic scene in a native 16:9 widescreen frame, photographed on a 50 mm anamorphic lens with motivated tungsten practical lighting, cool window fill, realistic skin texture, shallow depth of field, fine 35 mm grain, subtle halation and controlled highlight rolloff. No cuts; one continuous shot throughout. Describe the complete subject, wardrobe, environment, blocking, camera movement and spoken dialogue here.
overall_soundscape: Describe the synchronized ambience, physical sounds and nonverbal character sounds here.
non_diegetic_music: Describe the score and instrumentation here, or use N/A when the scene should rely only on natural sound.
```
## Important
The LoRA strengthens the visual cinematic language: framing, camera placement, perspective, lighting, lens feeling, production design and film texture. MiniMax H3 itself remains responsible for temporal motion, acting, speech, sound effects and music, so those elements should still be described clearly in the prompt.
For the best results, do not write only “cinematic.” Direct the scene like a filmmaker: specify the shot size, camera height, angle, lens, movement, light sources, subject blocking, environment, atmosphere and sound.