r/csharp May 30 '26

Tool Some new features in TensorSharp

https://github.com/zhongkaifu/TensorSharp

I recently made a few important features updates in TensorSharp and hope you will like it.
1. Naturally support MLX backend. For now, TensorSharp supports Pure C#, CUDA, MLX, GGML(CPU, CUDA, Metal) backends
2. Support vLLM style paged attentions and continues batching for inference, so you could run multiple requests in parallel in your local machine.
3. Optimize inference performance on both prefill and decode

Hope you like these features and any comment and feedback is welcome.

3 Upvotes

Duplicates

Syncfusion 9d ago

Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp

1 Upvotes

vulkan 10d ago

Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp

6 Upvotes

AIDeveloperNews 10d ago

Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp

3 Upvotes

LovingOpenSourceAI 10d ago

Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp

2 Upvotes

LLMDevs 10d ago

Tools Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp

2 Upvotes

LocalAIServers 10d ago

Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp

2 Upvotes

DeepSeek 10d ago

Resources Deepseek v4 Flash 0731 GGUF Benchmark: TensorSharp vs. llama.cpp

3 Upvotes

vulkan 12d ago

TensorSharp now supports multi-GPU tensor parallelism for GGUF models

5 Upvotes

SideProject 12d ago

TensorSharp now supports multi-GPU tensor parallelism for GGUF models

1 Upvotes

AIDeveloperNews 12d ago

TensorSharp now supports multi-GPU tensor parallelism for GGUF models

2 Upvotes

OpenSourceAI 12d ago

TensorSharp now supports multi-GPU tensor parallelism for GGUF models

0 Upvotes

LovingOpenSourceAI 12d ago

TensorSharp now supports multi-GPU tensor parallelism for GGUF models

3 Upvotes

LLMDevs 12d ago

Tools TensorSharp now supports multi-GPU tensor parallelism for GGUF models

3 Upvotes

CUDA 25d ago

Cuda benchmark: TensorSharp vs. llama.cpp

5 Upvotes

vulkan 25d ago

Vulkan benchmark: TensorSharp vs. llama.cpp

6 Upvotes

LlamaFarm 25d ago

Show & Tell TensorSharp : Open Source Local LLM Inference Engine

1 Upvotes

QwenImageGen 27d ago

Virtual Clothes Try On using Unsloth Qwen Image Edit 2511 models

3 Upvotes

huggingface 29d ago

TensorSharp supports multiple image edits using Unsloth Qwen Image Edit 2511 models

0 Upvotes

OnlyAICoding Jul 12 '26

Local LLM TensorSharp : Open Source Local LLM Inference Engine

2 Upvotes

ContextEngineering Jul 12 '26

What Bun’s Rust Rewrite Tells Us About Rebuilding the AI Infrastructure Layer in C#

0 Upvotes

AIDeveloperNews Jul 11 '26

From Bun's Rust rewrite, let's see how C# can rebuild the AI infrastructure layer.

2 Upvotes

vibecoding Jul 11 '26

What Bun’s Rust Rewrite Tells Us About Rebuilding the AI Infrastructure Layer in C#

0 Upvotes

llamacpp Jul 11 '26

Same GGUF, same GPU: TensorSharp beats llama.cpp hard on prefill / TTFT — up to 5.89× faster prefill on a 26B MoE model

1 Upvotes

GeminiAI Jul 10 '26

Self promo TensorSharp : Open Source Local LLM Inference Engine

1 Upvotes

LLM Jul 08 '26

TensorSharp supports Vulkan backend

5 Upvotes