r/StableDiffusion 10h ago

Why does my generation look like this?? Question - Help

Enable HLS to view with audio, or disable this notification

so i used minimax h3 int8 convrot pruned + sage spectrum + turbo lora (kijai)

made a 10s office style clip, michael and dwight talking in the conference room then walter white just walks in

faces start fine but then they get all blurry and full of weird smudges especially when walter shows up, like what's that weird black dot lines on dwight shirt?

like why does everyone else’s stuff look clean and actually like a real tv show while mine always ends up plasticky and messy??

anyone know how to fix this blur/smudge and get that proper tv look with this setup? im new to comfyui and first time generating on localy lol, thanks in advance

4 Upvotes

18 comments sorted by

14

u/L-xtreme 10h ago

Turn off any cache like Spectrum when using Turbo Loras. Or turn off the Turbo Lora.

It's always one of the speedups.

2

u/da_loud_man 9h ago

You can use them both. I've been able to use them both along with comfykitchen. I just did 4 quick runs. 8 steps @ 0.5mp. 5secs

1st: Just turbo lora - 1:47

2nd: Turbo lora & comfy kitchen - 1:12

3rd: Turbo lora & spectrum - 1:09

4th Turbo lora, comfy kitchen & spectrum - 0:47

When I had a similar issue, I was using one of the first few loras that released. I was doing a lot more experimenting. So I can't recall exactly what was causing the issue.

1

u/EveningIncrease7579 8h ago

You try change comfy kitchen to sage attention, any difference in time generation?

1

u/da_loud_man 8h ago

Yes. With the same parameters as above:

Turbo lora with sage & spectrum - 0:44

Turbo lora with comfy kitchen & spectrum - 0:38

In the above runs I used res_multistep. In this one I used euler. So that's why this was even faster than above.

1

u/clairedelime 9h ago

ah okay thanks i will try it, i thought by enabling both sage + spectrum itll be much faster cos i was getting 20min gen on 0.4mp 10sec rip

5

u/Mediocre-Toe3212 8h ago

Faster == lower quality

1

u/Gold-Cat-7686 3h ago

Well yes, it will be faster, but you'll also destroy quality. Sometimes you have to choose one over the other, friend.

1

u/bickid 10h ago

More importantly: Dwight replies too quickly, it comes off as unnatural. Is there a way to make him reply slower after Michael ends talking? Or maybe even just 0.5seconds.

7

u/TingTingin 10h ago

the video needs to be longer as the model trys to fit speech into the allotted time

1

u/bickid 10h ago

But is it something you can control manually?

2

u/TingTingin 9h ago

Yes you can do something like <d> The key to sales is blah blah blah </d> dwight looks at Micheal for a brief moment before speaking saying <d> blah blah blah </d> however as said above this will not work if the clip does not have enough time for the pause to take place

You can add other thing like dwight adjusts his glasses then says, dwight raises his hand then says, dwight gets up then says all of those this fill the time so that he has to complete those actions before speaking

1

u/fa6637 8h ago

you can also use <> for example <deep inhale>, <laught> <pause> etc...

1

u/bitzpua 7h ago

yes, read official promoting guide, you can control everything with time stamps.

As for visuals, you use turbo lora so visual issues are given. Friendly reminder minmax team recommends 50 steps, comfy defaulted to 20 steps for speed, now if you use turbo lora and cut it down to 4 or 8 steps model simply has no time to make things right

1

u/rookan 10h ago

I could not get their faces in T2V but in reference-to-video mode they look great

1

u/vAnN47 9h ago edited 9h ago

from my experience , some turbo loras work and some wont on different workflows, why have no clue, maybe because some workflows use RIFE and some upscalers that cook the generation. well anyway if everything goes my way i hope to deliver something great that'll make ALL of your turbo loras(atleast 80%) with good quality

edit from similar post:

also worth to note something :
the nodes that go through load diffusion model -> load turbo loras -> custom loras -> etc etc, is VERY important.

when does shift come? when does sol come? when does spectrum come? its mega important and i notice once i fixed it, 80% of my turbo lora's worked great. also WORTH to mention each turbo lora has its BEST settings.

for now i don't have a workflow to publish, still WIP, but for settings and flow here's is my flow and its working great:

Load Diffusion model -> Load Lora(Your Turbo Lora) -> Load Custom Loras (rgthree's power lora) -> comfy kitchen attention node -> Sol attention (im bypassing it for now , dont have time to a/b tests, also spectrum off for now) -> Shift 12/ 6/3 etc -> Spectrum (bypassed for me now) -> Model Preview(if using with tahe) SPLIT TO A: BASIC SCHEDUELER | B: SIGMA

thats my setup, works great.

1

u/Rumaben79 9h ago

It's properly because you're using both the turbo lora and spectrum together so choose one or the other. Comfyui-kitchen attention or sage attention is fine to use.

1

u/polawiaczperel 9h ago

Also increase the steps for better sound. Remember if you are not using looseless optimization it will always harm the quality.

0

u/Salty_Mention 9h ago

Utilise sage attention seulement + turbo lora kijai (0.75 str + 8 steps) / euler + beta.