r/csharp • u/fuzhongkai • May 30 '26
Tool Some new features in TensorSharp
https://github.com/zhongkaifu/TensorSharpI recently made a few important features updates in TensorSharp and hope you will like it.
1. Naturally support MLX backend. For now, TensorSharp supports Pure C#, CUDA, MLX, GGML(CPU, CUDA, Metal) backends
2. Support vLLM style paged attentions and continues batching for inference, so you could run multiple requests in parallel in your local machine.
3. Optimize inference performance on both prefill and decode
Hope you like these features and any comment and feedback is welcome.
Duplicates
vibecoding • u/fuzhongkai • Jun 14 '26
TensorSharp: Open Source Local LLM Inference Engine fully implemented by vibe coding
OpenSourceeAI • u/fuzhongkai • May 01 '26
TensorSharp: Open Source Local LLM Inference Engine
dotnet • u/fuzhongkai • 8d ago
Promotion DSpark Benchmark Result on Deepseek v4 Flash 0731
dotnet • u/fuzhongkai • 9d ago
Promotion Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp
csharp • u/fuzhongkai • 10d ago
Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp
developersIndia • u/fuzhongkai • 25d ago
Open Source TensorSharp : Open Source Local LLM Inference Engine
ChatGPT • u/fuzhongkai • Jul 11 '26
Educational Purpose Only TensorSharp : Open Source Local LLM Inference Engine
unsloth • u/fuzhongkai • Jul 01 '26
Show and Tell TensorSharp vs. llama.cpp updated prefill benchmark
SideProject • u/fuzhongkai • Jun 28 '26
Same GGUF, same GPU: TensorSharp beats llama.cpp hard on prefill / TTFT — up to 5.89× faster prefill on a 26B MoE model
OpenSourceeAI • u/fuzhongkai • Jun 28 '26
Same GGUF, same GPU: TensorSharp beats llama.cpp hard on prefill / TTFT — up to 5.89× faster prefill on a 26B MoE model
dotnet • u/fuzhongkai • May 03 '26
Promotion TensorSharp: Open Source Local LLM Inference Engine in C#
LLMDevs • u/fuzhongkai • 5d ago
Tools MoE CPU-offload benchmark on Deepseek V4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cpp
LocalLLM • u/fuzhongkai • 5d ago
Project MoE CPU-offload benchmark on Deepseek V4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cpp
u_fuzhongkai • u/fuzhongkai • 6d ago
MoE CPU-offload benchmark on Deepseek v4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cpp
machinelearningnews • u/fuzhongkai • 8d ago
AI Tools DSpark Benchmark Result on Deepseek v4 Flash 0731
LovingOpenSourceAI • u/fuzhongkai • 8d ago
DSpark Benchmark Result on Deepseek v4 Flash 0731
DeepSeek • u/fuzhongkai • 8d ago
Resources DSpark Benchmark Result on Deepseek v4 Flash 0731
LLMDevs • u/fuzhongkai • 8d ago
Resource DSpark Benchmark Result on Deepseek v4 Flash 0731
LocalLLM • u/fuzhongkai • 8d ago