🐰 Mello's Definitive Zootopia LoRA (Minimax H3) (Please Read) 🦊
Trained on 250 clips from "a certain animated film." It can do mostly any action you could want, except for nsfw stuff (that version is in development).
Full Showcase (used the older 64 rank version):
⚡ Quick Settings & Info
Trigger Word: z00t0p1a
Recommended Weight: 0.85 (Tested between 0.8 – 1.0, Motion becomes a little more stiff beyond .85)
Base Architecture: Minimax H3
Dataset: 250 video clips
Training Specs: 15,750 steps (~30 epochs)
📦 Version Guide: Which file should I download?
This model comes in two versions attached to the release:
🌟 Rank 128 (RECOMMENDED / FLAGSHIP):
This is the main version! Produces the most aesthetically pleasing renders, superior character accuracy, cleaner textures, and higher visual detail. Start with this one!⚡ Rank 64 (ALTERNATIVE):
Slightly more baked/overcooked, but offers higher motion dynamics/movements and slightly tighter audio/dialogue synchronization. If you like a prompt but the 128 doesn't seem to give quite the right result, try this one.
🔗 Links & Mirrors
💾 Direct Mirrors:
The workflow I used (Still using an old version, will change workflows at some point): Link
📋 Modular Prompt Reference Guide
💡 How to use this guide: Text inside (parentheses) is meant to be replaced with your specific details. Plain text outside parentheses can stay as-is, unless you combine with a style lora.
z00t0p1a, (Character 1 ex. Judy Hopps), (Character 2 ex. Nick Wilde, Pawbert, etc.). A (shot type ex. wide side-angle, medium close-up, high-angle tracking, over-the-shoulder) shot of (describe the main focus of the video ex. a grey 3D rendered anthropomorphic rabbit and a red 3D rendered anthropomorphic fox). (Describe their primary action, posture, and movement without detailing their specific anatomy yet ex. stand side-by-side inside a small white ZPD vehicle, runs frantically towards the right side of the camera, leans in close to the rabbit's face). (If a character is in frame and is wearing clothes or gear, describe it here ex. The grey rabbit is wearing a blue police uniform with a black bulletproof vest. The red fox is wearing a pink floral short-sleeved shirt with a patterned tie and light grey pants).
IF JUDY HOPPS (THE GREY RABBIT) IS IN FRAME, ADD THESE TO THE PROMPT, IF NOT SKIP:
The rabbit has detailed fluffy white and grey fur. It has a round head with two large ears and a small nose (Optional: "and a small puff tail poking out of her upper pants"). Her purple eyes are large and round and (Describe eye state and expression ex. wide open with a proud, content expression; half closed with a deeply exhausted, sleepy expression; slightly narrowed with a curious expression). The rabbit (Describe where and how the rabbit is looking ex. looks straight ahead at the fox's chest level, looks back over her shoulder, looks directly at the phone screen). She (Describe her mouth and expression ex. has a closed mouth with a subtle smile expression, is talking with a big cheerful open-mouthed smile, has a slightly open mouth with a gasping expression). (If the rabbit blinks, add "The rabbit blinks [multiple times/slowly]"). The rabbit's ears are (Describe ear position ex. up, down, pinned back, in a ball, blowing back in the wind).
IF NICK WILDE (THE RED FOX) IS IN FRAME, ADD THESE TO THE PROMPT, IF NOT SKIP:
The fox has detailed fluffy reddish-orange fur. It has a tapered head with two large, dark-tipped ears and a small black nose (Optional: and a large fluffy tail poking out of his upper pants). His green eyes are large and round and (Describe eye state and expression ex. half closed with a relaxed, affectionate gaze; wide open with a completely terrified, panicked expression; hidden behind reflective sunglasses). The fox (Describe where and how the fox is looking ex. looks down toward the rabbit, looks forward unconcerned of the background action, surveys his surroundings). He (Describe his mouth and expression ex. has a closed mouth with a subtle smile expression, is talking with a smug smile expression, has an open mouth baring his teeth). (If the fox blinks, add "The fox blinks [multiple times, once]"). The fox's ears are (Describe ear position ex. up, back, down, flattened back).
IF ANOTHER CHARACTER OR ANIMAL IS IN FRAME (EX. LYNX, SNAKE, BEAVER), APPLY THE SAME CHARACTERISTICS AS THE PREVIOUS SUBJECTS:
(Describe their fur/scales texture ex. The lynx has detailed, thick grey and white fur / The snake has detailed, shiny blue scales). (Describe their head shape, ears, and nose/fangs ex. It has a broad head with two large, pointed ears topped with black tufts and a small pink nose / It has a broad head with two large yellow eyes and one white fang). (Describe their eye color and state ex. His eyes are large and round and wide open with an eager expression / His yellow eyes with vertical pupils are half closed). (Describe their gaze direction). (Describe their mouth, teeth, and overall expression). (Note if they are blinking). (Describe their ear position if applicable).
IF A CHARACTER IS SPEAKING, ADD THIS SECTION (OPTIONAL):
The [character] says:
(0:00 - 0:02) (Add tone/action in parentheses if needed ex. Slowly/Sarcastically) "[Insert dialogue here]."
(0:03 - 0:05) "[Insert dialogue here]."
CAMERA, BACKGROUND, AND LIGHTING (ALWAYS INCLUDE AT THE END):
The camera (describe the camera movements here ex. is stationary with a subtle zoom in, tracks backwards rapidly with intense handheld shake to keep the running characters in frame, pans to follow the fast-moving car). The background is (describe the environment ex. a sunlit park field with soft-focus on a fountain, a dimly lit rustic wooden interior hallway, a bustling sun-drenched fish market pier). The lighting is (describe the lighting conditions ex. from soft daylight illuminating the stage evenly, from a bright warm offscreen setting sun casting long shadows, from cool blue ambient party lights). (Optional: Add post-processing effects ex. shallow depth of field, motion blur, vivid).🛠️ Tech-Help & Troubleshooting
"Speech is cutoff midway when generated video finishes"
If the dialogue between characters are cut off before the characters are supposed to finish, add-on to the prompt that a post-action happens or that there is a second of pause after they finish their conversation (honestly, it's my bad for not refining the dataset for minimax interms of speech but the workaround works for now).
Also, of note is if the voices don't quite match up in terms of sounding like them, calm down on either the vocalization styles and tone or if you are doing a stylized render, do one where it's not stylized on the same seed and sync up the speech after the fact (only voice guaranteed to work is judy, all the others do have a much higher chance of failing).
"Characters are too stiff"
Yeah, minimax hates adding unprompted actions unlike wan. If you do write out your own prompts, make sure that you use an llm to help add micro-movements and expressions into the prompts to give more life to the characters.
Just feeding the 'master prompt' into an llm won't give you great out of the box gens from my findings, you should have a stylistic idea or concept feed alongside the master prompt to give better and more unique results. Make sure to make use of involved camera movements ("Camera has handheld movements as it tracks with the subject", etc.).
"Character's faces/bodies are distorted/undefined at a distance"
This seems to be a problem with minimax in general. I did most of my gens at 1280x512 and ran into this issue quite a bit even when increasing steps.
Best recommendation I have is to just zoom in the shot from a wide to either medium-wide or medium. If you like that shot a lot however, rendering at 1920x804 (or its 1080p equivalent) does seem to help a lot, tho motion has the potential to suffer (I never found myself going this route since it would take over 45 mins to do 30 steps at that quality for me).
Some Other Notes:
Speed-Up Workflows: I never tested this with any speed-up stuff so whatever optimization you do to get loras to work with those workflows, just apply that to here.
Video Duration: Going beyond 12-13 seconds can be rough, especially 15 seconds, actions become repeated and/or stiff (the official dataset only goes up to 10 seconds with most being sub 6 second clips, mostly because it was made for wan).
Description
FAQ
Comments (16)
how did you train so many different characters so well? and yes, please do nsfw for science...animal science.
lol, Its a combination of a large dataset, detailed prompting for said dataset (adding detailed descriptions of background characters and such), low training rate, and a higher rank seems to help as well. I also duplicated some of the dataset videos for the lynx and the snake to help (since they are only a few videos of them in the set).
@mello_ai i said nude underneath her robe and it put a pink lingerie on her :( lol soon (lol forgot to remove that last reference to the lingerie you added, but i doubt that would help too )
@mello_ai plus i tried your same robe prompt without the lora, and yea this lora is a HUGE improvement over base H3.
@mello_ai actually I got it to kinda sorta work on this last post I made. always gives her big boobies though
@kronos1959777 yeah it can do rough nsfw stuff but it will need additional loras to get more desirable stuff (I posted an early nsfw version on my site if you'd like to grab it and test it: https://www.melloai.art/renders/models/zootopia_minimax/ (the combo model, also included rough documentation, its just not to my standards but might serve fine)
With the art style it probably helps to increase the rank to include that stuff too. But for example, all those paid movie loras we see, they could have all prompted characters too. The authors just generally avoid it, perhaps because they are paywalling them
@mello_ai lol i tried my own judy hopps lora i made before and it got her tits out but they look fake as shit.
omg bro
Does this work well with Ref2va?
I've only done a few tests but it seems to work ok (I attached two comedic gens that used a ref2va workflow at the end of my post showcase)
Let the games begin!
Zootopia is not on my list of things I want to gen, but the absolute quality of your showcase is indisputable. Best I've ever seen from H3
<3
Your model on Minimax is really good. I'm looking forward to your NSFW version of it. As well as seeing if you've prepared any other models.
💕👍🏻