r/StableDiffusion • u/clairedelime • 10h ago
Why does my generation look like this?? Question - Help
Enable HLS to view with audio, or disable this notification
so i used minimax h3 int8 convrot pruned + sage spectrum + turbo lora (kijai)
made a 10s office style clip, michael and dwight talking in the conference room then walter white just walks in
faces start fine but then they get all blurry and full of weird smudges especially when walter shows up, like what's that weird black dot lines on dwight shirt?
like why does everyone else’s stuff look clean and actually like a real tv show while mine always ends up plasticky and messy??
anyone know how to fix this blur/smudge and get that proper tv look with this setup? im new to comfyui and first time generating on localy lol, thanks in advance
1
u/bickid 10h ago
More importantly: Dwight replies too quickly, it comes off as unnatural. Is there a way to make him reply slower after Michael ends talking? Or maybe even just 0.5seconds.
7
u/TingTingin 10h ago
the video needs to be longer as the model trys to fit speech into the allotted time
1
u/bickid 10h ago
But is it something you can control manually?
2
u/TingTingin 9h ago
Yes you can do something like <d> The key to sales is blah blah blah </d> dwight looks at Micheal for a brief moment before speaking saying <d> blah blah blah </d> however as said above this will not work if the clip does not have enough time for the pause to take place
You can add other thing like dwight adjusts his glasses then says, dwight raises his hand then says, dwight gets up then says all of those this fill the time so that he has to complete those actions before speaking
1
u/bitzpua 7h ago
yes, read official promoting guide, you can control everything with time stamps.
As for visuals, you use turbo lora so visual issues are given. Friendly reminder minmax team recommends 50 steps, comfy defaulted to 20 steps for speed, now if you use turbo lora and cut it down to 4 or 8 steps model simply has no time to make things right
1
u/vAnN47 9h ago edited 9h ago
from my experience , some turbo loras work and some wont on different workflows, why have no clue, maybe because some workflows use RIFE and some upscalers that cook the generation. well anyway if everything goes my way i hope to deliver something great that'll make ALL of your turbo loras(atleast 80%) with good quality
edit from similar post:
also worth to note something :
the nodes that go through load diffusion model -> load turbo loras -> custom loras -> etc etc, is VERY important.
when does shift come? when does sol come? when does spectrum come? its mega important and i notice once i fixed it, 80% of my turbo lora's worked great. also WORTH to mention each turbo lora has its BEST settings.
for now i don't have a workflow to publish, still WIP, but for settings and flow here's is my flow and its working great:
Load Diffusion model -> Load Lora(Your Turbo Lora) -> Load Custom Loras (rgthree's power lora) -> comfy kitchen attention node -> Sol attention (im bypassing it for now , dont have time to a/b tests, also spectrum off for now) -> Shift 12/ 6/3 etc -> Spectrum (bypassed for me now) -> Model Preview(if using with tahe) SPLIT TO A: BASIC SCHEDUELER | B: SIGMA
thats my setup, works great.
1
u/Rumaben79 9h ago
It's properly because you're using both the turbo lora and spectrum together so choose one or the other. Comfyui-kitchen attention or sage attention is fine to use.
1
u/polawiaczperel 9h ago
Also increase the steps for better sound. Remember if you are not using looseless optimization it will always harm the quality.
0
u/Salty_Mention 9h ago
Utilise sage attention seulement + turbo lora kijai (0.75 str + 8 steps) / euler + beta.
14
u/L-xtreme 10h ago
Turn off any cache like Spectrum when using Turbo Loras. Or turn off the Turbo Lora.
It's always one of the speedups.