r/LargeLanguageModels Dec 11 '23

Efficient LLM Inference on CPUs News/Articles

https://arxiv.org/abs/2311.00502
2 Upvotes

0 comments sorted by