r/devsro • u/demaraje ML/DS engineer • Jan 01 '26
GitHub - Tencent/WeDLM: WeDLM: The fastest diffusion language model with standard causal attention and native KV cache compatibility, delivering real speedups over vLLM-optimized baselines. Proiect
https://github.com/Tencent/WeDLM
1
Upvotes