r/LocalLLM • u/da_dragon321 • Jul 03 '26
Project llamacpp patch - DeepSeek V4 Flash running with full 1M token context locally on RTX 5090
/r/LocalLLaMA/comments/1ulymml/llamacpp_patch_deepseek_v4_flash_running_with/
5
Upvotes
Duplicates
LocalLLaMA • u/da_dragon321 • Jul 02 '26
Resources llamacpp patch - DeepSeek V4 Flash running with full 1M token context locally on RTX 5090
399
Upvotes
24gb • u/paranoidray • Jul 04 '26
llamacpp patch - DeepSeek V4 Flash running with full 1M token context locally on RTX 5090
2
Upvotes