Sign In

CLEAN X2 E26 VIDEO VAE H3

Early Access

Updated: Sep 21, 2026

base model

Download

1 variant available

fp16 SafeTensor

H3_Clean_X2_E26.safetensors

Half precision, best balance • 4.89 GB

Verified:

Type
VAE
Stats

7

Reviews
Published

Sep 21, 2026

Base Model

MiniMax H3

Hash
AutoV2
216D7D71F3
default creator card background decoration
Followers - 1722

1.7K

Likes - 4757

4.8K

Downloads - 43725

43.7K

Rendered Romance Contest Participant

MiniMax H3 is licensed by MiniMax under the MiniMax H3 Community License Agreement. That agreement’s Applicable Territory excludes the European Union, the United Kingdom, the Republic of Korea and the United States of America. Your use of H3 and of any H3 derivative is subject to that agreement and its Acceptable Use Policy.

MiniMax H3

Clean X2 E26 video VAE is a 2× spatial video VAE for MiniMax H3, providing cleaner high-resolution decoding, improved detail preservation, and reduced grid artifacts.

Compatible with the X2 VAE decoding workflow.

This node package is required: https://github.com/TripleHeadedMonkey/ComfyUI-MiniMaxH3_LatentUpscaler

A little more information about CLEAN X2 E26 VIDEO VAE H3

No external upscaler.

The frames you're looking at are decoded directly from CLEAN X2 E26 VIDEO VAE H3 at 2× spatial resolution.

This project started because I wasn't satisfied with simply generating an H3 video and throwing an AI upscaler on top of it.

I wanted to know:

How much more can we get directly from H3's own decoder?

What followed was... a lot more work than I expected. 😅

Why is it called E26? 😀

E26 stands for the 26th major experimental iteration — but there were countless smaller experiments, training runs, failed ideas, A/B tests and intermediate versions inside those iterations, sleepless nights and the desire to give up 😅.

But I didn't give up, and here we are 😀👋🎉🥳🥹

What it does

-Native 2× VAE decoding

-No external AI upscaler

-No post-generation super-resolution model

-Clean output without the old repeating grid

-Works directly with real MiniMax H3 generations

-More spatial information than the normal decode