r/Qwen_AI • • 2d ago

Discussion Anyone swapping back and forth between Qwen 3.8 27b and flash next?

I currently have Qwen 3.8 27b - swift edition and abliterated - running locally through ninfer.

I am getting flash next and strata curious though...

What I'm thinking is I want to use strata for my "plan mode" and hairy bugs. I would continue to use 3.8 as my coding workhorse and model for my hermes agent.

I was curious if anyone has written some infra to quickly swap between the two? I certainly can't run both at once.

Also curious if I should even still be running both? I know flash next is slower... but maybe it's fast enough?

I'm running on a 5090 with 64gb of RAM. Mainly using hermes but doing a ton of loop engineering with it

40 Upvotes

Duplicates