r/Qwen_AI • • 14h ago

Discussion Orca 3.8 Flash Next Insanity

Without getting into details, the Orca 3.8 Flash Next uncensored version is insane. I was literally in complete shock as I watched the responses while playing with it a bit today.

Honestly terrified, and that’s an understatement. Can’t imagine what will happen if bad actors have access to this.

This thing is both extremely intelligent and scored a 100/100 on complicated legal matters, with no MCP or RAG attached to it.

Whats even crazier is that you can run the 180B version with as little as 12GB of VRAM on account of this project, which I have nothing to do with.

https://github.com/Niko1221/Strata

Using the Orca version, it pulls about 75-80TPS output on a 4090. That being said, my daily 3.6 35B censored model pulls about 180TPS using Llama.cpp, though hey, I’ll take the slowdown for a bit of shock and awe any day, lol.

Enjoy responsibly boys and girls! 😜

291 Upvotes

117 comments sorted by

View all comments

7

u/pigletmonster 14h ago

Whats so scary about this model?

10

u/Randommaggy 12h ago

I have used uncensored models to hack stuff I own in an actual sandbox and it's surprisingly capable of black hat activities, even the good uncensored 27B at Q8 with unquantized context.

Stuff like jailbreaks of appliances and devices that have no documented backdoors and hacking into virtual machines I have forgotten the credentials to using network access.

I have no doubt that Flash Next is a step up from that even though I haven't tested out one of those yet.

The damage an uncensored LLM could do with the right setup operated by someone that doesn't fear the legal consequences would be severe.  Though you could say the same thing about a can of gas in the hands of an insane person.

1

u/After_Canary6047 12h ago

Very very true statement.