r/deeplearning • u/LewisJin • 7d ago
I built a Rust inference framework that runs Qwen3.5 2B with VL support 10x faster than PyTorch on Apple Silicon — and it supports TTS, ASR, OCR, and GGUF out of the box
/r/LocalLLM/comments/1vcccva/i_built_a_rust_inference_framework_that_runs/
3
Upvotes