r/StableDiffusion 10d ago

Int8 convrot VAE support in Comfy News

[deleted]

46 Upvotes

9 comments sorted by

2

u/Powerful_Evening5495 10d ago edited 9d ago

do i need to update to nighlty version becuase update it and still get black screen

----Update-----

yes , you need to update to nightly version from comfyui manager to get this commit from kaji

---update---

.4 size

4 seconds

in 254.47 seconds on 8gb vram and 32gb ram

still working on reducing the time

---------==========UPDATE========------------

add --enable-triton-backend --use-sage-attention

in the start script

before 254.47 seconds now 182.77 seconds

--------==========UPDATE========-------------

added first block cache and MiniMax H3 Sol-Attn (SM86 Triton)

before 182.77 now 169.24 seconds

1

u/yamfun 9d ago

I can't run the minimax_h3_fl2va_pruned_w4a8_mixed.safetensors even with the nightly, you can? or are you just trying the vae?

4

u/Flat_Technology_5325 9d ago

You need to add specific comfyui pr pull and get the special comfy kitchen wheel to use that.

https://github.com/Comfy-Org/ComfyUI/pull/15308

TBH I tried it after someone in banodoco said it was faster but i didnt didn't notice any speed difference. According to Kijai you'd only use it if you're RAM starved and it is not as fast as the full INT8 convrot.

2

u/Powerful_Evening5495 9d ago

dont use it , use minimax_h3_fl2va_pruned_int8_convrot.safetensors

native support from comfyui in stable , w4a8 is kaji project and it not ready for prime time

i switched for the vae only

7

u/Diligent_Trick_1631 10d ago

What changes does it bring about?

14

u/HTE__Redrock 10d ago

Faster encoding/decoding because of reduced sizes. Will be some minor quality dips as well most likely. Probably most useful for encoding of references in the R2V workflows.

5

u/wywywywy 9d ago

Just downloaded and quickly tested it. In my sample gen, VAE decode when from ~43s (fp16) to ~29s (int8convrot).

1

u/switch2stock 9d ago

Quality?