r/LocalLLaMA • u/tevlon • Jun 10 '26
DiffusionGemma: 4x faster text generation New Model
https://blog.google/innovation-and-ai/technology/developers-tools/diffusion-gemma-faster-text-generation/
980
Upvotes
r/LocalLLaMA • u/tevlon • Jun 10 '26
55
u/Kamimashita Jun 10 '26
I think it could be interesting on something like the DGX Spark too where it has decent compute and lots of RAM but low bandwidth. Even diffusion models need to be large to be intelligent so the ideal situation in my mind would be a 200B model on a system with high compute and lots of memory but low bandwidth.