Download
2 variants available
bf16 SafeTensor
Zootopia_v7_SFW_128_Public_Release.safetensors
BF16, good balance • 1.16 GB
Verified: 2 days ago
SafeTensor
2580 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
160 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
1.3K0 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9K
(29)
Sep 13, 2026
MiniMax H3

630 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
700 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
MiniMax H3 is licensed by MiniMax under the MiniMax H3 Community License Agreement. That agreement’s Applicable Territory excludes the European Union, the United Kingdom, the Republic of Korea and the United States of America. Your use of H3 and of any H3 derivative is subject to that agreement and its Acceptable Use Policy.
MiniMax H3
🐰 Mello's Definitive Zootopia LoRA (Minimax H3) (Please Read) 🦊
Trained on 250 clips from "a certain animated film." It can do mostly any action you could want, except for nsfw stuff (that version is in development).
Full Showcase (used the older 64 rank version):
⚡ Quick Settings & Info
Trigger Word: z00t0p1a
Recommended Weight: 0.85 (Tested between 0.8 – 1.0, Motion becomes a little more stiff beyond .85)
Base Architecture: Minimax H3
Dataset: 250 video clips
Training Specs: 15,750 steps (~30 epochs)
📦 Version Guide: Which file should I download?
This model comes in two versions attached to the release:
🌟 Rank 128 (RECOMMENDED / FLAGSHIP):
This is the main version! Produces the most aesthetically pleasing renders, superior character accuracy, cleaner textures, and higher visual detail. Start with this one!⚡ Rank 64 (ALTERNATIVE):
Slightly more baked/overcooked, but offers higher motion dynamics/movements and slightly tighter audio/dialogue synchronization. If you like a prompt but the 128 doesn't seem to give quite the right result, try this one.
🔗 Links & Mirrors
💾 Direct Mirrors:
The workflow I used (Still using an old version, will change workflows at some point): Link
📋 Modular Prompt Reference Guide
💡 How to use this guide: Text inside (parentheses) is meant to be replaced with your specific details. Plain text outside parentheses can stay as-is, unless you combine with a style lora.
z00t0p1a, (Character 1 ex. Judy Hopps), (Character 2 ex. Nick Wilde, Pawbert, etc.). A (shot type ex. wide side-angle, medium close-up, high-angle tracking, over-the-shoulder) shot of (describe the main focus of the video ex. a grey 3D rendered anthropomorphic rabbit and a red 3D rendered anthropomorphic fox). (Describe their primary action, posture, and movement without detailing their specific anatomy yet ex. stand side-by-side inside a small white ZPD vehicle, runs frantically towards the right side of the camera, leans in close to the rabbit's face). (If a character is in frame and is wearing clothes or gear, describe it here ex. The grey rabbit is wearing a blue police uniform with a black bulletproof vest. The red fox is wearing a pink floral short-sleeved shirt with a patterned tie and light grey pants).
IF JUDY HOPPS (THE GREY RABBIT) IS IN FRAME, ADD THESE TO THE PROMPT, IF NOT SKIP:
The rabbit has detailed fluffy white and grey fur. It has a round head with two large ears and a small nose (Optional: "and a small puff tail poking out of her upper pants"). Her purple eyes are large and round and (Describe eye state and expression ex. wide open with a proud, content expression; half closed with a deeply exhausted, sleepy expression; slightly narrowed with a curious expression). The rabbit (Describe where and how the rabbit is looking ex. looks straight ahead at the fox's chest level, looks back over her shoulder, looks directly at the phone screen). She (Describe her mouth and expression ex. has a closed mouth with a subtle smile expression, is talking with a big cheerful open-mouthed smile, has a slightly open mouth with a gasping expression). (If the rabbit blinks, add "The rabbit blinks [multiple times/slowly]"). The rabbit's ears are (Describe ear position ex. up, down, pinned back, in a ball, blowing back in the wind).
IF NICK WILDE (THE RED FOX) IS IN FRAME, ADD THESE TO THE PROMPT, IF NOT SKIP:
The fox has detailed fluffy reddish-orange fur. It has a tapered head with two large, dark-tipped ears and a small black nose (Optional: and a large fluffy tail poking out of his upper pants). His green eyes are large and round and (Describe eye state and expression ex. half closed with a relaxed, affectionate gaze; wide open with a completely terrified, panicked expression; hidden behind reflective sunglasses). The fox (Describe where and how the fox is looking ex. looks down toward the rabbit, looks forward unconcerned of the background action, surveys his surroundings). He (Describe his mouth and expression ex. has a closed mouth with a subtle smile expression, is talking with a smug smile expression, has an open mouth baring his teeth). (If the fox blinks, add "The fox blinks [multiple times, once]"). The fox's ears are (Describe ear position ex. up, back, down, flattened back).
IF ANOTHER CHARACTER OR ANIMAL IS IN FRAME (EX. LYNX, SNAKE, BEAVER), APPLY THE SAME CHARACTERISTICS AS THE PREVIOUS SUBJECTS:
(Describe their fur/scales texture ex. The lynx has detailed, thick grey and white fur / The snake has detailed, shiny blue scales). (Describe their head shape, ears, and nose/fangs ex. It has a broad head with two large, pointed ears topped with black tufts and a small pink nose / It has a broad head with two large yellow eyes and one white fang). (Describe their eye color and state ex. His eyes are large and round and wide open with an eager expression / His yellow eyes with vertical pupils are half closed). (Describe their gaze direction). (Describe their mouth, teeth, and overall expression). (Note if they are blinking). (Describe their ear position if applicable).
IF A CHARACTER IS SPEAKING, ADD THIS SECTION (OPTIONAL):
The [character] says:
(0:00 - 0:02) (Add tone/action in parentheses if needed ex. Slowly/Sarcastically) "[Insert dialogue here]."
(0:03 - 0:05) "[Insert dialogue here]."
CAMERA, BACKGROUND, AND LIGHTING (ALWAYS INCLUDE AT THE END):
The camera (describe the camera movements here ex. is stationary with a subtle zoom in, tracks backwards rapidly with intense handheld shake to keep the running characters in frame, pans to follow the fast-moving car). The background is (describe the environment ex. a sunlit park field with soft-focus on a fountain, a dimly lit rustic wooden interior hallway, a bustling sun-drenched fish market pier). The lighting is (describe the lighting conditions ex. from soft daylight illuminating the stage evenly, from a bright warm offscreen setting sun casting long shadows, from cool blue ambient party lights). (Optional: Add post-processing effects ex. shallow depth of field, motion blur, vivid).🛠️ Tech-Help & Troubleshooting
"Speech is cutoff midway when generated video finishes"
If the dialogue between characters are cut off before the characters are supposed to finish, add-on to the prompt that a post-action happens or that there is a second of pause after they finish their conversation (honestly, it's my bad for not refining the dataset for minimax interms of speech but the workaround works for now).
Also, of note is if the voices don't quite match up in terms of sounding like them, calm down on either the vocalization styles and tone or if you are doing a stylized render, do one where it's not stylized on the same seed and sync up the speech after the fact (only voice guaranteed to work is judy, all the others do have a much higher chance of failing).
"Characters are too stiff"
Yeah, minimax hates adding unprompted actions unlike wan. If you do write out your own prompts, make sure that you use an llm to help add micro-movements and expressions into the prompts to give more life to the characters.
Just feeding the 'master prompt' into an llm won't give you great out of the box gens from my findings, you should have a stylistic idea or concept feed alongside the master prompt to give better and more unique results. Make sure to make use of involved camera movements ("Camera has handheld movements as it tracks with the subject", etc.).
"Character's faces/bodies are distorted/undefined at a distance"
This seems to be a problem with minimax in general. I did most of my gens at 1280x512 and ran into this issue quite a bit even when increasing steps.
Best recommendation I have is to just zoom in the shot from a wide to either medium-wide or medium. If you like that shot a lot however, rendering at 1920x804 (or its 1080p equivalent) does seem to help a lot, tho motion has the potential to suffer (I never found myself going this route since it would take over 45 mins to do 30 steps at that quality for me).
Some Other Notes:
Speed-Up Workflows: I never tested this with any speed-up stuff so whatever optimization you do to get loras to work with those workflows, just apply that to here.
Video Duration: Going beyond 12-13 seconds can be rough, especially 15 seconds, actions become repeated and/or stiff (the official dataset only goes up to 10 seconds with most being sub 6 second clips, mostly because it was made for wan).

.jpeg)