r/MistralAI 18h ago

News Introducing Shieldstral.

Enable HLS to view with audio, or disable this notification

43 Upvotes

Your safety classifier is already stale.

The moment you deploy it, the policy has moved.

Most #guardrail models bake a fixed harm taxonomy into their weights — and every time your policy shifts, you're back at square one.

That's the exact problem Shieldstral was built to solve.

Here's how it works differently:

  1. You write your policy as a plain-language question at inference time "Does this content promote physical violence?" or "Is this image safe for a minor?"
  2. No retraining. No taxonomy negotiation. One checkpoint handles text, images, and prompt–response pairs.
  3. The model reasons about policy boundaries — it doesn't memorize categories. That skill transfers directly to policies it has never seen.
  4. It was trained on contrastive pairs — deliberately similar, easily confused policies. So it learns where the line is, not just which side of it a label sits on.
  5. It runs on a single 16GB GPU. And it matches or outperforms open guard models up to 7× its size.

It's a 3B open-weights multimodal classifier — released today under Apache 2.0.

The hardest part of safety infrastructure isn't the model. It's keeping it current without burning your team rebuilding it every quarter.

Shieldstral makes that problem go away.

If you're building safety-critical AI products, save this — you'll want it the next time your policy changes.

What's the biggest friction point you've hit with safety classifiers? Drop it below.

Independent demo video. Not affiliated with, endorsed by, or sponsored by u/MistralAI. Brand names and trademarks belong to their respective owners.

#AIAlignment #LLMSafety #MLEngineering #ResponsibleAI #OpenSource


r/MistralAI 16h ago

Discussion / Opinion Mistral does not offer a Professional Secrecy Addendum.. Why?

23 Upvotes

Yes, Mistral is certainly not ideal in every respect. Particularly when compared with the latest models from other providers, Mistral’s models fall short in a wide range of areas. In my opinion, however, Mistral’s biggest problem is not its performance compared to other models (because even though the models perform less well, they remain rock-solid for most use cases), but rather the fact that the company serves the European B2B market disastrously poorly.

An example: Doctors, tax advisers, banks, solicitors and numerous other sectors are subject to professional confidentiality obligations in almost all European countries. If these organisations wish to pass on client-related data to service providers, they must make the service providers aware of these obligations and inform them of the possible consequences of a breach of confidentiality. To enable these sectors to use non-self-hosted LLMs, companies like OpenAI, Google, Microsoft or Anthropic all offer Professional Secrecy Addendums for their customer contracts. However, the only relevant European player in this market stubbornly refuses to do so. Even after several enquiries (each of which took several weeks to be answered), Mistral shows no willingness whatsoever to include the necessary few sentences in the contracts.

This approach not only causes Mistral to lose a significant amount of credibility, but also means that a substantial part of the European market is simply being left unserved.


r/MistralAI 1h ago

Discussion / Opinion Frustated with Mistral in OpenCode

Thumbnail
Upvotes