Shieldstral

Shieldstral

Mistral AI
+
+

Related Products

  • Dialpad Support
    1,588 Ratings
    Visit Website
  • BAND
    3 Ratings
    Visit Website
  • Josys
    246 Ratings
    Visit Website
  • Skillcast
    1,106 Ratings
    Visit Website
  • Pipefy
    592 Ratings
    Visit Website
  • cside
    37 Ratings
    Visit Website
  • Pikmykid
    235 Ratings
    Visit Website
  • Auth0
    1,067 Ratings
    Visit Website
  • Coevera
    752 Ratings
    Visit Website
  • Addigy
    261 Ratings
    Visit Website

About

Zero out your backlog and increase precision with automated Trust and Safety decisions — no eng, ML, or vendor management required. Define what you do and don’t want on your platform — our AI agents will execute those policies with better-than-human precision and speed. Our quality monitoring tools let you deploy correct, repeatable policies with confidence. Legacy ML models require expensive labeling and training, and training human reviewers takes even longer. We allow you to deploy new enforcements today with no agent or model retraining. It’s impossible to understand why ML classifiers make a prediction, and human reviewers aren’t much better. SafetyKit provides detailed reasoning with every decision.

About

Shieldstral is a 3B open-weights, policy-adaptive multimodal safety classifier designed to evaluate text, images, and text-plus-image content using policies defined at inference time. Instead of relying on a fixed taxonomy of harm categories, it frames moderation as a binary question-answering task: users provide an instruction describing the evaluation context and strictness, a yes-or-no safety question, and the content to judge. The model reads the “yes” and “no” logits and converts them into a continuous, calibrated safety score, allowing applications to threshold or rank results by confidence rather than depend on a single discrete label. This formulation unifies prompt classification, response moderation, refusal detection, toxicity detection, and multimodal safety in one interface, while letting teams adapt policies without retraining the model. Shieldstral can evaluate prompts, responses, prompt-response pairs, images, and images with accompanying text.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Businesses looking for a powerful Content Moderation platform

Audience

AI platform teams that need customizable, multimodal content moderation without retraining a separate safety model for every policy

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

No information available.
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

SafetyKit
www.getsafetykit.com

Company Information

Mistral AI
Founded: 2023
France
mistral.ai/news/shieldstral/

Alternatives

Shieldstral

Shieldstral

Mistral AI

Alternatives

Intrinsic

Intrinsic

Decoy Technologies

Categories

Categories

Integrations

Mistral AI
Velt

Integrations

Mistral AI
Velt
Claim SafetyKit and update features and information
Claim SafetyKit and update features and information
Claim Shieldstral and update features and information
Claim Shieldstral and update features and information