CivArchive
    CLEAN X2 E26 VIDEO VAE H3 - v1.0

    Clean X2 E26 video VAE is a 2× spatial video VAE for MiniMax H3, providing cleaner high-resolution decoding, improved detail preservation, and reduced grid artifacts.

    Compatible with the X2 VAE decoding workflow.

    This node package is required: https://github.com/TripleHeadedMonkey/ComfyUI-MiniMaxH3_LatentUpscaler

    A little more information about CLEAN X2 E26 VIDEO VAE H3

    No external upscaler.

    The frames you're looking at are decoded directly from CLEAN X2 E26 VIDEO VAE H3 at 2× spatial resolution.

    This project started because I wasn't satisfied with simply generating an H3 video and throwing an AI upscaler on top of it.

    I wanted to know:

    How much more can we get directly from H3's own decoder?

    What followed was... a lot more work than I expected. 😅

    Why is it called E26? 😀

    E26 stands for the 26th major experimental iteration — but there were countless smaller experiments, training runs, failed ideas, A/B tests and intermediate versions inside those iterations, sleepless nights and the desire to give up 😅.

    But I didn't give up, and here we are 😀👋🎉🥳🥹

    What it does

    -Native 2× VAE decoding

    -No external AI upscaler

    -No post-generation super-resolution model

    -Clean output without the old repeating grid

    -Works directly with real MiniMax H3 generations

    -More spatial information than the normal decode

    Description

    Clean X2 E26 video VAE is a 2× spatial video VAE for MiniMax H3, providing cleaner high-resolution decoding, improved detail preservation, and reduced grid artifacts.

    Compatible with the X2 VAE decoding workflow.

    This node package is required: https://github.com/TripleHeadedMonkey/ComfyUI-MiniMaxH3_LatentUpscaler

    A little more information about CLEAN X2 E26 VIDEO VAE H3

    No external upscaler.

    The frames you're looking at are decoded directly from CLEAN X2 E26 VIDEO VAE H3 at 2× spatial resolution.

    This project started because I wasn't satisfied with simply generating an H3 video and throwing an AI upscaler on top of it.

    I wanted to know:

    How much more can we get directly from H3's own decoder?

    What followed was... a lot more work than I expected. 😅

    Why is it called E26? 😀

    E26 stands for the 26th major experimental iteration — but there were countless smaller experiments, training runs, failed ideas, A/B tests and intermediate versions inside those iterations, sleepless nights and the desire to give up 😅.

    But I didn't give up, and here we are 😀👋🎉🥳🥹

    What it does

    -Native 2× VAE decoding

    -No external AI upscaler

    -No post-generation super-resolution model

    -Clean output without the old repeating grid

    -Works directly with real MiniMax H3 generations

    -More spatial information than the normal decode

    And will try to improve it

    VAE
    MiniMax H3

    Details

    Downloads
    6
    Platform
    CivitAI
    Platform Status
    Available
    Created
    9/21/2026
    Updated
    9/22/2026
    Deleted
    -

    Files

    cleanX2E26VIDEOVAEH3_v10.safetensors

    Mirrors