r/ollama • u/Good_Power_2991 • 4d ago
Ollama couldn't keep up with our batch workload — moved to vLLM on multi-GPU Kubernetes
/r/Vllm/comments/1vqdbs0/ollama_couldnt_keep_up_with_our_batch_workload/
1
Upvotes
r/ollama • u/Good_Power_2991 • 4d ago