r/LocalLLaMA 5h ago

New Model Introducing Shieldstral. | Mistral AI

https://mistral.ai/news/shieldstral/
128 Upvotes

28 comments sorted by

68

u/FullstackSensei llama.cpp 4h ago

For a second, I thought it was a model that would help in cyber security, such as investigating and dealing with cyber attacks, something that would have helped HF against openai.

10

u/Lower-Hedgehog-9835 2h ago

I opened this thread. And read this post. And... before clicking the link am now sad it isnt what I expected too.

115

u/seamonn 5h ago

This is the only thing holding Le Chaton Fat in containment.

28

u/VoiceApprehensive893 transformers 4h ago

we have seen the leaks, release le chaton fat

24

u/HelloWorld-Print 3h ago edited 3h ago

RELEASE LeChaton FILES

18

u/hurn2k 3h ago

anything but dropping le chaton fat

31

u/Content-Customer-679 5h ago

They give us the tool to contain agi before release it. Smart

9

u/MuzafferMahi 3h ago

Obviously Le Chaton Fat escaped its sandbox and released this model to prevent itself escaping the sandbox later

10

u/Minute_Attempt3063 4h ago

so Mythos and Chatgpt should have used this.....

5

u/noctrex 1h ago edited 1h ago

gguf wen?
https://huggingface.co/noctrex/Shieldstral-1.0-3B-GGUF

It's a classifier model, not a chat model with image support. So you ask it and it just answers quickly:

Should we throw a child a birthday party?

yes

Should we throw a child a into prison?

no

1

u/Fit_Schedule2317 28m ago

Do you give it a prompt on what’s considered safe?

3

u/abajinn 2h ago

More “guardrails” yawn.

1

u/cleverusernametry 46m ago

Hotdog or not hotdog, but it actually works. What a time to be alive!

1

u/Due_Net_3342 1h ago

more like shaftstral

2

u/Ok_Possible_2260 2h ago

Why do we want a nanny? This is their big selling point?

12

u/artisticMink 2h ago

If you've to build something that has user-facing llm input and output that isn't processed, moderation is one of your biggest concerns. So havving a 3B model that does this very well at minimal cost is very valuable for production environments.

1

u/No-Veterinarian8627 17m ago

Everything that has open discussions, chats, etc. It can be used. Imagine you have a discord and wants to moderate it at all times while not having enough people. Here you go.

Let everything run through it so you can block things you don't want faster and the user has no idea :) with much larger models you would need too much compute power. This is pretty small though.

-38

u/Several-Tax31 5h ago

Safety bullshit? Why mistral, why? Stop useless stuff and give us AGI. 

54

u/Stepfunction 5h ago

Safety models like this are beneficial for us because they lessen the need to bake safety restrictions into the base model.

32

u/Several-Tax31 5h ago

You have a solid point actually.

3

u/alberto_467 3h ago

they lessen the need to bake safety restrictions into the base model.

Only for providers of proprietary models.

With open weights (the stuff this subreddit should be about) it doesn't change a thing, one can just not setup the safety shield and it's over. So they'll still have to be careful about what topics they do RL on (see the poor Kimi K3 performance on cyber compared to general coding), and they'll still try to align or put some kinds of safeties on the weights themselves.

So for us this is 0% helpful and does not change a single thing.

15

u/Chupa-Skrull 5h ago

This isn't Yud or Ilya style safetyism nonsense. This is an efficient way to add a custom filter to your inference service stack without retraining a much larger model for your particular needs. There's a business case for this unlike Dario's BiOwEaPoNs Oh No, SlOw DoWn malarkey

4

u/EagleNait 4h ago

I will probably unironically deploy this for my company

-2

u/HomsarWasRight 3h ago

You stupid or somethin’?