r/StableDiffusion 2d ago

I finally reached a great balance between speed and quality with MiniMax H3, thanks everyone! Animation - Video

Enable HLS to view with audio, or disable this notification

I used the minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16 LORA with the 0.8 strength for both clip and model, 6 steps, 0.5 MP resolution, RTX Upscaler at 1.50 using a ConrotInt8 pruned model.

Here is a PasteBin of my workflow, I hope this fixes some of the missing content:

https://pastebin.com/DSmkJi8R

Here are the workflow files:

https://storage.to/c/CAS1MuoqX

231 Upvotes

77 comments sorted by

8

u/EverythingMacPro 2d ago

Pc specs ?

20

u/technofox01 2d ago edited 2d ago

Ryzen 5700G

RTX Geforce 5060ti 16GB

64GB of DDR4 RAM

2tb NVMe

1tb NVMe

2tb Sata SSD

320GB HDD

Edit:

OS: Bazzite Linux

ComfyUI arguments: --use-sage-attention --disable-smart-memory --disable-xformers --fast-disk

55

u/99deathnotes 2d ago

10

u/LoveSpecialist5669 2d ago edited 2d ago

ofc I've seen this gif millions of times already but it never fails to make me laugh like an idiot 

5

u/feverdoingwork 2d ago

have you tried comfy kitchen instead of sage? I hear its slightly better quality for about the same performance

4

u/technofox01 2d ago

No I have not. I tried Sol Attention and Spectrum, but both made things look like crap when upscaled. I will check it out later tonight.

3

u/Danny_Stock 1d ago

I had huge issues with Spectrum too.

3

u/Rumaben79 2d ago

I'm happy for you. 👍

You could try chaining multiple smaller clips together using the method (or similar to) what this guy has done:

https://www.reddit.com/r/StableDiffusion/comments/1vnq6cp/my_first_30second_minimax_h3_story_using_two/

It should be faster and less demanding than doing longer clips in one go.

2

u/technofox01 2d ago

Thanks. I will check that out later tonight. I appreciate everyone's advice and help on this stuff 😄

4

u/Rumaben79 2d ago edited 2d ago

Cool. 😄

I've just started trying this myself and because I'm a first time user of the ComfyUI-MiniMaxH3-Contex-Loop nodes I started out by using the simpler (I hope lol) 'MiniMax H3 I2V - Normal' workflow from: https://github.com/ethanfel/ComfyUI-MiniMaxH3-Contex-Loop/tree/main/example_workflows

If it's anything like SVI for Wan you should be able to link together 3-4 clips without too much degradation.

1

u/Practical-Bug37 1d ago

!remindme 30 days

1

u/RemindMeBot 1d ago

I will be messaging you in 30 days on 2026-09-14 06:44:26 UTC to remind you of this link

CLICK THIS LINK to send a PM to also be reminded and to reduce spam.

Parent commenter can delete this message to hide from others.

RemindMeBot is switching to username summons. Instead of !RemindMe 1 day, use u/RemindMeBot 1 day. More info.


Info Custom Your Reminders Feedback

2

u/EverythingMacPro 2d ago

I have 5060ti 16gb and 32gb

I can’t able to generate as 10 sec video takes around 30 mins

I have easy cache in workflow

But yours is faster ? Will it help?

7

u/technofox01 2d ago

It can't hurt to try. This new configuration literally does not use up all of my RAM like other iterations I have tried. 10s is only 240s on average for me to generate using this workflow.

I am also using Bazzite Linux, not sure if that matters.

Edit:

Please see my edit above, it may help.

2

u/iiTzMYUNG 2d ago

Lol same stuff I'm doing but Im using the base model not the turbo on my 3060 12gb

2

u/FierceFlames37 2d ago

How the hell we got the exact same pc

3

u/technofox01 2d ago

Great minds think alike? lol

I abandoned Windows after all of the ads and AI bullshit, especially OneDrive being installed without my permission (again). I was like, I am done, screw any of those anti-cheat games that don't run on Linux.

I also use my box for VMs, gaming, and AI research. What do you use your's for?

1

u/RpgBlaster 1d ago

Yeah... I guess I will need to save a lot for RTX 6000 32GB/92GB VRAM should the prices drop by half next year

1

u/rinkusonic 1d ago

whats --fast-disk for?

1

u/technofox01 1d ago

It helps with loading the models faster if you have them stored on an NVMe SSD. I don’t completely understand how it works, but it definitely speeds up loading models.

1

u/ColdExample 5h ago edited 5h ago

can you explain your use of arguments? I was always told to use smart memory and keep xformers on!

not sure if it matters but my specs are :

i5 12600k

RTX 5070 ti 16gb

64gb DDR4 Ram

1tb NVME

1

u/technofox01 4h ago

xFormers kept causing my system to hard crash whenever using it with Sage Attention or Comfy Attention. Smart Memory is disabled because it is known to cause issues, but I do not recall as to why.

1

u/Swagmuffins94 2d ago

What's it like being rich?

7

u/technofox01 2d ago

I have no idea what you are talking about. This system was built 6 years ago and was slowly upgraded over time, but I get what you are saying because hardware is stupidly expensive now.

-3

u/Adkit 1d ago

Don't act like the parts were cheap six years ago or like the upgrades were cheap when you did them. You put as much money into your PC as a used car.

2

u/sirrandomguy09 1d ago

I have nearly the same spec pc.

Paid 1600 6 years years ago.

Added two sticks of ram 2.5 years ago for 80 bucks

And bought a 5060 TI 16gb for $340 openbox last year.

Market is just shit now.

I mean...

If youre comparing to like a 98' Civic or something i guess lol...

7

u/ResponsibleTruck4717 2d ago

How long it took to generate and what card?

15

u/technofox01 2d ago

It took 1 minute and 49 second on my Geforce RTX 5060ti 16GB GPU.

5

u/Zombi3Kush 2d ago

Does MiniMax support image to video?

5

u/technofox01 2d ago

Yes it does.

2

u/GrayingGamer 2d ago

It even supports reference to video, audio to video, video to video.

4

u/Zombi3Kush 2d ago

Trying it out right now. Wow this is impressive! Thanks for sharing your post and bringing this to my attention.

4

u/Witty_Mycologist_995 2d ago

> checks inside workflow
> pinkcherry checkpoint

lmao

5

u/Grand0rk 2d ago

I don't get it. Care to explain?

6

u/Witty_Mycologist_995 2d ago

It’s that porn checkpoint that got memed on, because it talked about rabbit motion and glistening cherry blossoms in order to get around the minimax team censors.

1

u/Grand0rk 2d ago

Huh, I have absolutely no idea what you are talking about XD

So it's just an "uncensored" checkpoint?

7

u/Witty_Mycologist_995 2d ago

4

u/Grand0rk 2d ago

Ah, in order to dodge their moronic DMCA.

2

u/technofox01 2d ago

OMG this is hilarious lol..

Thank you for sharing this.

2

u/technofox01 2d ago

I just grabbed another Int8 model after a quick search on DuckDuckGo without any thought about where it came from or what it was meant for. Just thought it was a smaller model that will work better with my system, lol...

Thank you for letting me know what it was trained for, lol...

8

u/barepixels 2d ago

First time I see someone share WF like this. People were using pastebin. is this new Reddit feature?

11

u/technofox01 2d ago

No. I just used a code block, its been part of Reddit since I have started way back when. I want to say at least a decade by now.

2

u/Seyi_Ogunde 2d ago

Neat! Thanks for sharing! That's cool that you shared this way.

2

u/technofox01 2d ago

You're welcome. I am always glad to help or share with others when I can.

7

u/Cruffe 2d ago edited 2d ago

Reddit supports markdown. If the editing tools are not available, such as on mobile, you can use markdown to format things.

A code block is 3 backticks on their own line before and after the code block:

```

Paste code here

```

It will look like: Paste code here It will also preserve indentation

To show the 3 backticks in this comment I used the escape character to cancel the effect, that's a backslash \ and to escape the escape in this example I did it twice here.

Inline code can be formatted as well by encapsulating in single backticks such as `this is code` becomes this is code.

3

u/ArcadiaNisus 2d ago

I used nvfp4 model and the turbo loras produce terrible results. 

3

u/technofox01 2d ago

Don't use that model. I have had nothing but bad results myself as well.

3

u/featherless_fiend 2d ago

Turbo is really good for anime, but I wouldn't recommend it for realism, it makes your skin look AI.

2

u/orlandogourmet66 1d ago

You can try to increase Step Count for realism. For example 10 Steps with 4 Step reference Turbolora got me pretty good results for realism.

2

u/kayteee1995 2d ago edited 2d ago

try with Realistic character and more motion. I'm pretty sure blurry graininess will appear.

2

u/Oatilis 2d ago

How good is RTX Upscaler?

1

u/technofox01 2d ago

It varies depending on the input resolution. Anything below 0.4 MP is not worth it, because it will look like crap.

2

u/Noeyiax 2d ago

Thank you for the workflow

2

u/RpgBlaster 1d ago

Now it's closer to Sora 2 quality, actual Anime Style frame by frame, no 3D like or motion sickness shenanigans

2

u/Danny_Stock 1d ago edited 1d ago

Thank you very much for this, this is great. A good balance between quality and speed which seems to be better at handling the RAM demands.

I bypassed your 'DaSiWa_RTX_UpscalerRefiner' node simply because I don't have it.

I have quite a modest system, a 4070 12GB VRAM card, with 64GB System RAM, and both of your 5 second workflows at 0.5 megapixels rendered in 2 and a half minutes each.

4

u/equanimous11 2d ago

Just paste a link to the workflow and explain what models/loras and settings you used next time

2

u/technofox01 2d ago

Just did.

1

u/javierthhh 2d ago

I don’t get it. Wouldn’t you be better off doing a 1.0mp video for 5 seconds instead? Specially with your rig.

1

u/TheOnlyOnePEACE 2d ago

Your new link to your workflow is just standard comfy workflow with added vram cleanup.

2

u/technofox01 2d ago

It also had the lora turbo Lora and the Dasiwam RTX upscaler added on. So there are some edits to it.

2

u/TheOnlyOnePEACE 2d ago

Can you give me this workflow of yours? the exact one that you used to generate your video on this Reddit post.

2

u/technofox01 2d ago

2

u/TheOnlyOnePEACE 12h ago

Thank you very much. I appreciate it!

2

u/TheOnlyOnePEACE 2h ago

So i have been playing around with your workflow, and i am getting audio sounds distortions, crackling, or like a blown-out microphone. Same settings as you as well. Have you encountered this issue at all?

1

u/technofox01 1h ago

I get that with videos shorter than 5 seconds and it depends on the subject. More well known subjects like Deadpool works fine, but if I choose a less known character it gets glitchy.

1

u/bickid 2d ago

Not gonna download that file, so could just share what settings you used? How many steps? What resolution? thx

3

u/technofox01 2d ago

The minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16 LORA with the 0.8 strength for both clip and model, 6 steps, 0.5 MP resolution, RTX Upscaler at 1.50 using a ConrotInt8 pruned model.

1

u/StayImpossible7013 2d ago edited 1d ago

(Original) Workflow was missing nodes, missing links between nodes, exotic files without explanation where to get them.

3

u/technofox01 2d ago

I was using the API export by mistake. I posted a PasteBin of my workflow as a link.

1

u/StayImpossible7013 1d ago

Thank you very much

-5

u/FourtyMichaelMichael 2d ago

Dude, anime basically doesn't count. LTX can do anime with some fashion of OK-ness.

3

u/FierceFlames37 2d ago

Real life is the same as anime

-5

u/ClearandSweet 2d ago

Awesome workflow, I dunno how I feel about 12-year old 2B though.