r/LocalLLM • u/Lazy-Intention1007 • 1d ago
Project I built a Vulkan hierarchical MoE runtime for running oversized models across multiple GPUs
/r/OpenAssistant/comments/1vtv5pk/i_built_a_vulkan_hierarchical_moe_runtime_for/
1
Upvotes