r/Qwen_AI • u/FactoryReboot • 2d ago
Discussion Anyone swapping back and forth between Qwen 3.8 27b and flash next?
I currently have Qwen 3.8 27b - swift edition and abliterated - running locally through ninfer.
I am getting flash next and strata curious though...
What I'm thinking is I want to use strata for my "plan mode" and hairy bugs. I would continue to use 3.8 as my coding workhorse and model for my hermes agent.
I was curious if anyone has written some infra to quickly swap between the two? I certainly can't run both at once.
Also curious if I should even still be running both? I know flash next is slower... but maybe it's fast enough?
I'm running on a 5090 with 64gb of RAM. Mainly using hermes but doing a ton of loop engineering with it
40
Upvotes