MiniMax-H3-VAE-ONNX

ONNX version of the MiniMax-H3 VAE for ComfyUI, which can increase speed by up to 1.7x

preivew

Usages

  1. Put encoder & decoder .onnx files (and .data files, if present) into ComfyUI/models/vae

    You can use the w4a16_awq decoder if you don't have 12GB+ of VRAM.

  2. Install this extension: ComfyUI-H3VAE_TRT

Benchmark

Tested at 1344×768 @ 5s empty latent:

Model Name Size Dec Time Speedup Rel RMS Error PSNR (dB) Has NaN
fp16_baseline 4.85 GB 17.87s 1.00x 0.000% Inf (Exact) False
trt_fp16_engine 4.85 GB 11.84s 1.50x 0.117% 65.84 dB False
trt_w4a16_engine 1.54 GB 13.10s 1.36x 6.778% 30.65 dB False

License

The weights in this repository are released under the Apache 2.0 License. Use of the MiniMax-H3 base model is also subject to its corresponding license and terms of use.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for lihaoyun6/MiniMax-H3-VAE-ONNX

Quantized
(10)
this model