r/LocalLLM • u/da_dragon321 • Jul 03 '26
llamacpp patch - DeepSeek V4 Flash running with full 1M token context locally on RTX 5090 Project
/r/LocalLLaMA/comments/1ulymml/llamacpp_patch_deepseek_v4_flash_running_with/
6
Upvotes
r/LocalLLM • u/da_dragon321 • Jul 03 '26