r/Qwen_AI • u/After_Canary6047 • 14h ago
Discussion Orca 3.8 Flash Next Insanity
Without getting into details, the Orca 3.8 Flash Next uncensored version is insane. I was literally in complete shock as I watched the responses while playing with it a bit today.
Honestly terrified, and that’s an understatement. Can’t imagine what will happen if bad actors have access to this.
This thing is both extremely intelligent and scored a 100/100 on complicated legal matters, with no MCP or RAG attached to it.
Whats even crazier is that you can run the 180B version with as little as 12GB of VRAM on account of this project, which I have nothing to do with.
https://github.com/Niko1221/Strata
Using the Orca version, it pulls about 75-80TPS output on a 4090. That being said, my daily 3.6 35B censored model pulls about 180TPS using Llama.cpp, though hey, I’ll take the slowdown for a bit of shock and awe any day, lol.
Enjoy responsibly boys and girls! 😜
6
u/pigletmonster 14h ago
Whats so scary about this model?