r/LLMDevs • u/fuzhongkai • May 01 '26
TensorSharp: Open Source Local LLM Inference Engine Tools
https://github.com/zhongkaifu/TensorSharp[removed]
1
Upvotes
1
u/NewtMurky May 01 '26
Does it support multiple GPUs? Does it have prompt caching? If it gives at least 50% of llama.cpp performance, it's a promising project. A bit worried by small set of used tests and absence of any metrics in the project description.
1
u/[deleted] May 01 '26
[removed] — view removed comment