Shieldstral

Shieldstral

Mistral AI
+
+

Related Products

  • SDS Manager
    4 Ratings
    Visit Website
  • Parasoft
    148 Ratings
    Visit Website
  • HSI Donesafe
    176 Ratings
    Visit Website
  • Macaw AMS
    8 Ratings
    Visit Website
  • SciSure
    299 Ratings
    Visit Website
  • Intelex
    176 Ratings
    Visit Website
  • MaintainX
    2,605 Ratings
    Visit Website
  • PackageX OCR Scanning
    48 Ratings
    Visit Website
  • Astra Pentest
    283 Ratings
    Visit Website
  • ERA EHS Software
    47 Ratings
    Visit Website

About

Preamble's AI Safety and Security Platform is an integrated solution designed to streamline and enhance the management of AI systems within an organization. It offers a centralized hub for managing people, overseeing diverse data labeling projects, providing clear guidelines for consistent data labeling, and tracking all labels and datasets. The platform also facilitates the evaluating of custom models and serves as a comprehensive center for AI safety and security testing and policy deployment. From real-time engagement with AI models to rigorous policy testing, the platform combines these multifaceted components to ensure alignment with organizational values, ethical principles, and compliance standards. Whether it's managing individual roles, conducting adversarial testing, or deploying safety controls, Preamble's platform offers a cohesive and user-friendly environment that addresses the complex and evolving needs of AI safety and security.

About

Shieldstral is a 3B open-weights, policy-adaptive multimodal safety classifier designed to evaluate text, images, and text-plus-image content using policies defined at inference time. Instead of relying on a fixed taxonomy of harm categories, it frames moderation as a binary question-answering task: users provide an instruction describing the evaluation context and strictness, a yes-or-no safety question, and the content to judge. The model reads the “yes” and “no” logits and converts them into a continuous, calibrated safety score, allowing applications to threshold or rank results by confidence rather than depend on a single discrete label. This formulation unifies prompt classification, response moderation, refusal detection, toxicity detection, and multimodal safety in one interface, while letting teams adapt policies without retraining the model. Shieldstral can evaluate prompts, responses, prompt-response pairs, images, and images with accompanying text.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Any company using generative AI and LLMs.

Audience

AI platform teams that need customizable, multimodal content moderation without retraining a separate safety model for every policy

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

$100/month/user
Custom prices available for enterprises.
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Preamble
Founded: 2021
United States
www.preamble.com

Company Information

Mistral AI
Founded: 2023
France
mistral.ai/news/shieldstral/

Alternatives

Alternatives

Shieldstral

Shieldstral

Mistral AI

Categories

Categories

Content Moderation Features

Artificial Intelligence
Audio Moderation
Brand Moderation
Comment Moderation
Customizable Filters
Image Moderation
Moderation by Humans
Reporting / Analytics
Social Media Moderation
User-Generated Content (UGC) Moderation
Video Moderation

Integrations

ChatGPT
Claude
Cohere
IBM watsonx.data
Llama 2
Mistral AI
OpenAI
Reddit

Integrations

ChatGPT
Claude
Cohere
IBM watsonx.data
Llama 2
Mistral AI
OpenAI
Reddit
Claim Preamble and update features and information
Claim Preamble and update features and information
Claim Shieldstral and update features and information
Claim Shieldstral and update features and information