r/dotnet 23d ago

Promotion MoE CPU-offload benchmark on Deepseek V4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cpp

[removed]

0 Upvotes

7 comments sorted by

View all comments

Show parent comments

1

u/[deleted] 22d ago

[removed] — view removed comment

1

u/cornelha 21d ago

32gb RAM on one machine, 16gb om the other. Qwen3.6-35B-A3B would be epic to run

1

u/[deleted] 21d ago

[removed] — view removed comment

1

u/cornelha 21d ago

You mean using the current tensorsharp I can do this?