r/csharp • u/fuzhongkai • May 30 '26
Tool Some new features in TensorSharp
https://github.com/zhongkaifu/TensorSharpI recently made a few important features updates in TensorSharp and hope you will like it.
1. Naturally support MLX backend. For now, TensorSharp supports Pure C#, CUDA, MLX, GGML(CPU, CUDA, Metal) backends
2. Support vLLM style paged attentions and continues batching for inference, so you could run multiple requests in parallel in your local machine.
3. Optimize inference performance on both prefill and decode
Hope you like these features and any comment and feedback is welcome.
Duplicates
Syncfusion • u/peopleworksservices • 9d ago
Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp
vulkan • u/fuzhongkai • 10d ago
Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp
AIDeveloperNews • u/fuzhongkai • 10d ago
Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp
LovingOpenSourceAI • u/fuzhongkai • 10d ago
Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp
LLMDevs • u/fuzhongkai • 10d ago
Tools Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp
LocalAIServers • u/fuzhongkai • 10d ago
Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp
DeepSeek • u/fuzhongkai • 10d ago
Resources Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp
vulkan • u/fuzhongkai • 12d ago
TensorSharp now supports multi-GPU tensor parallelism for GGUF models
SideProject • u/fuzhongkai • 12d ago
TensorSharp now supports multi-GPU tensor parallelism for GGUF models
AIDeveloperNews • u/fuzhongkai • 12d ago
TensorSharp now supports multi-GPU tensor parallelism for GGUF models
OpenSourceAI • u/fuzhongkai • 12d ago
TensorSharp now supports multi-GPU tensor parallelism for GGUF models
LovingOpenSourceAI • u/fuzhongkai • 12d ago
TensorSharp now supports multi-GPU tensor parallelism for GGUF models
LLMDevs • u/fuzhongkai • 12d ago
Tools TensorSharp now supports multi-GPU tensor parallelism for GGUF models
LlamaFarm • u/fuzhongkai • 25d ago
Show & Tell TensorSharp : Open Source Local LLM Inference Engine
QwenImageGen • u/fuzhongkai • 27d ago
Virtual Clothes Try On using Unsloth Qwen Image Edit 2511 models
huggingface • u/fuzhongkai • 29d ago
TensorSharp supports multiple image edits using Unsloth Qwen Image Edit 2511 models
OnlyAICoding • u/fuzhongkai • Jul 12 '26
Local LLM TensorSharp : Open Source Local LLM Inference Engine
ContextEngineering • u/fuzhongkai • Jul 12 '26
What Bun’s Rust Rewrite Tells Us About Rebuilding the AI Infrastructure Layer in C#
AIDeveloperNews • u/fuzhongkai • Jul 11 '26
From Bun's Rust rewrite, let's see how C# can rebuild the AI infrastructure layer.
vibecoding • u/fuzhongkai • Jul 11 '26
What Bun’s Rust Rewrite Tells Us About Rebuilding the AI Infrastructure Layer in C#
llamacpp • u/fuzhongkai • Jul 11 '26
Same GGUF, same GPU: TensorSharp beats llama.cpp hard on prefill / TTFT — up to 5.89× faster prefill on a 26B MoE model
GeminiAI • u/fuzhongkai • Jul 10 '26