Clean X2 E26 video VAE is a 2× spatial video VAE for MiniMax H3, providing cleaner high-resolution decoding, improved detail preservation, and reduced grid artifacts.
Compatible with the X2 VAE decoding workflow.
This node package is required: https://github.com/TripleHeadedMonkey/ComfyUI-MiniMaxH3_LatentUpscaler
A little more information about CLEAN X2 E26 VIDEO VAE H3
No external upscaler.
The frames you're looking at are decoded directly from CLEAN X2 E26 VIDEO VAE H3 at 2× spatial resolution.
This project started because I wasn't satisfied with simply generating an H3 video and throwing an AI upscaler on top of it.
I wanted to know:
How much more can we get directly from H3's own decoder?
What followed was... a lot more work than I expected. 😅
Why is it called E26? 😀
E26 stands for the 26th major experimental iteration — but there were countless smaller experiments, training runs, failed ideas, A/B tests and intermediate versions inside those iterations, sleepless nights and the desire to give up 😅.
But I didn't give up, and here we are 😀👋🎉🥳🥹
What it does
-Native 2× VAE decoding
-No external AI upscaler
-No post-generation super-resolution model
-Clean output without the old repeating grid
-Works directly with real MiniMax H3 generations
-More spatial information than the normal decode
Description
Clean X2 E26 video VAE is a 2× spatial video VAE for MiniMax H3, providing cleaner high-resolution decoding, improved detail preservation, and reduced grid artifacts.
Compatible with the X2 VAE decoding workflow.
This node package is required: https://github.com/TripleHeadedMonkey/ComfyUI-MiniMaxH3_LatentUpscaler
A little more information about CLEAN X2 E26 VIDEO VAE H3
No external upscaler.
The frames you're looking at are decoded directly from CLEAN X2 E26 VIDEO VAE H3 at 2× spatial resolution.
This project started because I wasn't satisfied with simply generating an H3 video and throwing an AI upscaler on top of it.
I wanted to know:
How much more can we get directly from H3's own decoder?
What followed was... a lot more work than I expected. 😅
Why is it called E26? 😀
E26 stands for the 26th major experimental iteration — but there were countless smaller experiments, training runs, failed ideas, A/B tests and intermediate versions inside those iterations, sleepless nights and the desire to give up 😅.
But I didn't give up, and here we are 😀👋🎉🥳🥹
What it does
-Native 2× VAE decoding
-No external AI upscaler
-No post-generation super-resolution model
-Clean output without the old repeating grid
-Works directly with real MiniMax H3 generations
-More spatial information than the normal decode
And will try to improve it