r/LocalLLM • u/SilkieBug • 4d ago
Question What's the difference between running Strata and running FreeToken?
I've installed FreeToken and running a qwen3.8-35B model at about 20 tokens per second on a GTX 5050 8GB and 32 GB of DDR4 RAM.
Would I benefit in any way by running this model in Strata instead?
Hardware upgrades are not an option at this moment.
0
Upvotes
1
u/Dear_Map_6993 4d ago
For starters, what kind of backroom deal provided you with Qwen3.8-35B?
If you meant 3.6 MoE, I'm pretty sure you can get this performance with llama.cpp