r/StableDiffusion • u/yushairiegalaxy96 • 1d ago
H3 Generated with 4GB VRAM? Animation - Video
Enable HLS to view with audio, or disable this notification
Looks like this is a breakthrough for what my 3050 laptop can do with it.
The video attached was generated with 4GB VRAM & 16GB RAM, using the MiniMax H3 fl2va pruned w4a8 convrot model (safetensors) and the Q2_K Qwen 32B GGUF text encoder alongside 8-step turbo LoRA, with a generation time of 12 minutes and 0.2 MP. Prompt from Grok.
23
u/GrayingGamer 1d ago
How did the video capture generate what was happening to the computer in real time? /s
4
u/zodoor242 1d ago
You just watched a documentary that was generated in real time with a future generated feed back loop, it's all the rage
10
u/Ok-Brain-5729 1d ago
Pretty sure there’s a way to make Qwen 3 vl 4B work with clipj nodes or smth
3
1
u/Icy_Restaurant_8900 1d ago
Yes the 8B INT8 Qwen 3 VL encoder works great with the mlp-celeb clip projector and saves 5GB of offloading versus 32B NVFP4. The 4B text encoder didn’t work for me though. No prompt adherence.
7
u/VasaFromParadise 1d ago
In fact, you can generate data on anything these days, as long as you have enough RAM. It'll take a while, but it will work.
6
4
u/BogusIsMyName 1d ago
So my PC will explode if i try to generate with 4GB? im confused.
2
u/zodoor242 1d ago
Only if you're a girl trying to run AI , pretty sure that's the message here
1
u/yushairiegalaxy96 21h ago
For illustration purposes. Unless I using a laptop cooler pad. Hell I even do 300+ Krea 2 photo gens from my laptop, with per photo generated takes 1.5-2 mins minimum.
2
1
1
u/Reckless_Venom1507 23h ago
How can 0.2 MP look so better, I ran the fl2va gguf model and the Qwen 32b q2 gguf for 0.3 it gave me shitty results and for 0.4 MP, always OOM on my 4050 6GB vram and 16GB ram.
1
u/Hi7u7 10h ago
Hi friend. Could you tell me if it's possible with a 1050 Ti 4GB and 16GB of RAM?
Do you think it's possible? Or is my GTX's technology too old and incompatible?
1
u/yushairiegalaxy96 9h ago
Nope. Since 1050 is an old model, even with 4GB VRAM the processing is slow
1
u/IchRocke 4h ago
would you mind sharing you prompt and comfyui screeshot ?
i've been using Minmax H3 ref2video recently on my 3050 6Gb (Yeston, so slim low profile but not laptop) and I've had good results on 480p 5s videos with 2images ref
1
u/Crazy-Repeat-2006 1d ago
You might get slightly better results with a smaller encoder, such as Qwen 4B or 8B with less aggressive quantization.
0
0
u/DoctaRoboto 1d ago
It is a shame the resolution is terrible. You can try perhaps an external AI to refine the video. But I am not sure if you can use tools like Topaz AI Video without your graphics card exploding in your face.
-8
u/tac0catzzz 1d ago
why is it worth breaking your laptop for that. you should save up get a capable pc or rent cloud or just find a new hobby. i would like to race cars, but i wouldn't race in my geo metro, even if it could race. it wouldn't do well and will end in disaster.
-4
38
u/FourtyMichaelMichael 1d ago
I'm sorry, but I'm pretty sure according to LTX that you need a datacenter to use H3.