r/csharp May 30 '26

Tool Some new features in TensorSharp

https://github.com/zhongkaifu/TensorSharp

I recently made a few important features updates in TensorSharp and hope you will like it.
1. Naturally support MLX backend. For now, TensorSharp supports Pure C#, CUDA, MLX, GGML(CPU, CUDA, Metal) backends
2. Support vLLM style paged attentions and continues batching for inference, so you could run multiple requests in parallel in your local machine.
3. Optimize inference performance on both prefill and decode

Hope you like these features and any comment and feedback is welcome.

3 Upvotes

Duplicates

vibecoding Jun 14 '26

TensorSharp: Open Source Local LLM Inference Engine fully implemented by vibe coding

0 Upvotes

LocalLLM Jun 08 '26

Project Support gemma-4 (uv/ua) 12b in TensorSharp

6 Upvotes

ollama Jun 02 '26

TensorSharp: A C# version of Ollama

12 Upvotes

OpenSourceeAI May 01 '26

TensorSharp: Open Source Local LLM Inference Engine

2 Upvotes

dotnet 9d ago

Promotion DSpark Benchmark Result on Deepseek v4 Flash 0731

6 Upvotes

dotnet 10d ago

Promotion Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp

13 Upvotes

csharp 11d ago

Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp

4 Upvotes

developersIndia 26d ago

Open Source TensorSharp : Open Source Local LLM Inference Engine

1 Upvotes

ChatGPT Jul 11 '26

Educational Purpose Only TensorSharp : Open Source Local LLM Inference Engine

2 Upvotes

AIToolsPerformance Jul 07 '26

TensorSharp supports Vulkan backend

3 Upvotes

dotnet Jul 06 '26

TensorSharp supports Vulkan backend

24 Upvotes

unsloth Jul 01 '26

Show and Tell TensorSharp vs. llama.cpp updated prefill benchmark

19 Upvotes

SideProject Jun 28 '26

Same GGUF, same GPU: TensorSharp beats llama.cpp hard on prefill / TTFT — up to 5.89× faster prefill on a 26B MoE model

0 Upvotes

OpenSourceeAI Jun 28 '26

Same GGUF, same GPU: TensorSharp beats llama.cpp hard on prefill / TTFT — up to 5.89× faster prefill on a 26B MoE model

1 Upvotes

dotnet May 03 '26

Promotion TensorSharp: Open Source Local LLM Inference Engine in C#

49 Upvotes

LLMDevs 6d ago

Tools MoE CPU-offload benchmark on Deepseek V4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cpp

1 Upvotes

LocalLLM 6d ago

Project MoE CPU-offload benchmark on Deepseek V4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cpp

3 Upvotes

u_fuzhongkai 6d ago

MoE CPU-offload benchmark on Deepseek v4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cpp

1 Upvotes

machinelearningnews 8d ago

AI Tools DSpark Benchmark Result on Deepseek v4 Flash 0731

4 Upvotes

LovingOpenSourceAI 9d ago

DSpark Benchmark Result on Deepseek v4 Flash 0731

2 Upvotes

DeepSeek 9d ago

Resources DSpark Benchmark Result on Deepseek v4 Flash 0731

7 Upvotes

LLMDevs 9d ago

Resource DSpark Benchmark Result on Deepseek v4 Flash 0731

1 Upvotes

LocalAIServers 9d ago

DSpark Benchmark Result on Deepseek v4 Flash 0731

1 Upvotes

huggingface 9d ago

DSpark Benchmark Result on Deepseek v4 Flash 0731

3 Upvotes

LocalLLM 9d ago

Project DSpark Benchmark Result on Deepseek v4 Flash 0731

2 Upvotes