If you think you can use 4 step mp only
The stage 2 is a turbo refinement and if you use that by it self you get a audio that sound like robot deepfring it self in hot oil. also you will get a overly contrasted image and many bugs
Dont use turbo in my opnion
How It Works
Instead of brute-forcing a high-resolution generation in one pass, this pipeline divides the workload:
Stage 1 (Motion & Composition): The workflow begins by generating a 0.3-megapixel base sample for 25 steps. By restricting the pixel density, the model is forced to focus entirely on physics and movement. This prevents the AI from hallucinating extra fingers or breaking anatomy during fast motion.
Stage 2 (Turbo Refinement): The stable base is passed into a latent upscale at 0.8 megapixels, followed by a highly efficient 4-step turbo pass. The turbo model skips the heavy lifting of calculating motion vectors and acts purely as a powerful denoiser, stripping away the initial grain and smearing while injecting crisp, high-res details.
Key Features & Benefits
Artifact-Free Fast Motion: Completely eliminates notorious MiniMax smearing and anatomical glitches.
Hardware Efficient: Highly optimized for local execution. It runs beautifully on standard consumer cards (like a 12GB RTX 4070) without triggering out-of-memory (OOM) errors.
Wan-Style Architecture: Brings high-noise to low-noise cascade logic to MiniMax H3.
Description
FAQ
Comments (2)
Hi, is this only meant to improve low resolution generations or is it also helpfull when someone is already generationg at 1mp or above? Thanks!
If you doing 1 mp or more you need to incress the mp in the
(2sampleupscale engine(12Gvram optimized))
from 0.8 mp to somthing like 1.3 mp
It will work