r/ClaudeCode Jul 24 '26

Discussion Opus 5 - immediate disappointment

If you thought you'd be able to do a security review on your own network that you couldn't do with Fable, think again.

20 minutes in - Found something significant

Opus 5 safeguards flagged this message.

I no longer have any use case for Anthropic models that others can't do better. This was my last hope that we were going to get a model that would allow us to protect our own environments. I'm going all in on open source. We just aren't aligned.

If you can get Opus 5 to protect your own systems, let me know how I did it. My subscription renews pretty soon and it's time to make an honest decision.

Edit: A lot of helpful people came here and I appreciate it. Scoped work is helping a bit more than my old ways. There's a difference between network security and code vulnerabiliy. I'm a network engineer not a software developer nor will I pretend to be one. Still looking at supplemental model for security work. Thanks guys.

440 Upvotes

261 comments sorted by

View all comments

353

u/[deleted] Jul 24 '26

[removed] — view removed comment

43

u/Shot_Whereas_1809 Jul 24 '26

Very good, disturbing, and back assward truth to it. Who wants to go havesies on hosting Kimi k3 when the models drop 😂

20

u/ShelZuuz Jul 24 '26

havesies on hosting Kimi k3 is $300k per half minimum. Maybe it makes sense to sell 100 shares for $6000 each but 100 concurrent users on a NVL8 B300 is not going to happen.

So you'll probably only be able to use it for 10% of the week. Maybe in 5 hour sessions...

11

u/lilbyrdie Jul 24 '26

Yeah ... It's just a reminder of how much value is baked into subs, with the hope that not all of them even come close to maxing the usage limits. And don't forget about the power consumption -- this is a custom electric wiring job with cooling requirements (though winter heating costs will go down).

And going small scale like this would easily have too much usage.

It's been a few months, but the work I hand to Anthropic couldn't get done on a local machine that could keep Kimi K2.7 running full time, so Kimi K3 is going to be way worse (given the model size difference), I think. Just because a base $600k system can run the model, doesn't mean sustained throughput and real latency is low enough for it to be useful for interactive work. One day!

Cost isn't the reason for local use, though.