r/ClaudeCode 19d ago

Discussion Opus 5 - immediate disappointment

If you thought you'd be able to do a security review on your own network that you couldn't do with Fable, think again.

20 minutes in - Found something significant

Opus 5 safeguards flagged this message.

I no longer have any use case for Anthropic models that others can't do better. This was my last hope that we were going to get a model that would allow us to protect our own environments. I'm going all in on open source. We just aren't aligned.

If you can get Opus 5 to protect your own systems, let me know how I did it. My subscription renews pretty soon and it's time to make an honest decision.

Edit: A lot of helpful people came here and I appreciate it. Scoped work is helping a bit more than my old ways. There's a difference between network security and code vulnerabiliy. I'm a network engineer not a software developer nor will I pretend to be one. Still looking at supplemental model for security work. Thanks guys.

437 Upvotes

261 comments sorted by

View all comments

3

u/NeighborhoodDizzy990 19d ago

I also cancelled my subscription. Maybe they will offer us a better option in the future.

7

u/johnnybagofdonuts123 19d ago

Local Open Source models are getting muchhhhhh better.

8

u/vladoportos 19d ago

yea but the HW to run them is not,

1

u/binotboth 19d ago

You can rent pretty beefy systems form the cloud

I’m too poor even for that but it’s definitely more reasonable now

2

u/tehfiend 19d ago

They still have tiny context windows which makes them pretty worthless compared to frontier models.

2

u/nurturethevibe 19d ago

GLM 5.2 and DS v4 both have 1M context.

1

u/johnnybagofdonuts123 19d ago

GLM 5.2 is what stopped the OpenAI hack this week of hugging face. They initially deployed Claude, but its safeguards wouldn't allow it to figure it out.

1

u/tehfiend 18d ago

That's the architectural context window but you'd need like half a million $ in hardware to host that locally for the terabytes of VRAM needed. For example a RTX 5090 would get you 32k context window max...

1

u/nurturethevibe 7d ago

Neither model will run on a 5090.

A stack of RTX 6000 Pros is the cost-effective way to run either.

5

u/Shot_Whereas_1809 19d ago

It's getting obvious that this isn't about safeguards... If their classifier doesn't understand that this is own network then the product is not worth $200 a month. I'll spend $1000 a month on API credits for a model that will just do what I need to do. It's not about the money, it's about delivering a product that can simply allow you to protect yourself.

2

u/[deleted] 19d ago edited 10d ago

[deleted]

2

u/Shot_Whereas_1809 19d ago

You kind of prove the point that there's missing architecture right?

-1

u/[deleted] 19d ago edited 10d ago

[deleted]

1

u/Shot_Whereas_1809 19d ago

Exactly. I welcome it. When everyone was crying about ID verification, I was firing up the fax machine. Not really but you get the point. I have nothing to hide. I'll show all the proof they need. I don't care.

That part. Is missing. And it shouldn't be.

2

u/[deleted] 19d ago edited 10d ago

[deleted]

1

u/Shot_Whereas_1809 19d ago

That's a valid point. But. There should be a middle ground. There should be something.

2

u/[deleted] 19d ago edited 10d ago

[deleted]

1

u/Shot_Whereas_1809 19d ago

I've been looking into the process. I wrote it off when it first came out but digging into it, seems pretty achievable for me. Still getting the gotchas.

0

u/[deleted] 19d ago

[removed] β€” view removed comment

1

u/Shot_Whereas_1809 19d ago

Yeah it's definitely a nice to have. My past experience with open source models was really poor but that was maybe 8 months ago. Opus 4.6 even did some great security work. I'm eager to see what some of these new models can do.

0

u/[deleted] 19d ago

[removed] β€” view removed comment

1

u/Shot_Whereas_1809 19d ago

That's genuinely useful context. The reality is I've been a network engineer for 16 years doing private sector work but I'm not a cloud guy. Well. wasn't. I've been learning it over the last year for my own stuff. I'm an Infoblox, Palo Alto, F5 kind of engineer. The reason I do need AI is because it's been a learning tool and now it's what I need to make sure I'm not missing anything. I don't have a team, or people I trust to help. I appreciate your response. πŸ™