Shieldstral

Shieldstral

Mistral AI
+
+

Related Products

  • TrafficGuard
    89 Ratings
    Visit Website
  • SciSure
    299 Ratings
    Visit Website
  • Feroot
    32 Ratings
    Visit Website
  • Curtain MonGuard Screen Watermark
    7 Ratings
    Visit Website
  • Retool
    593 Ratings
    Visit Website
  • StackAI
    53 Ratings
    Visit Website
  • Runpod
    220 Ratings
    Visit Website
  • Google AI Studio
    30 Ratings
    Visit Website
  • Gemini Enterprise Agent Platform
    999 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website

About

Llama Guard is an open-source safeguard model developed by Meta AI to enhance the safety of large language models in human-AI conversations. It functions as an input-output filter, classifying both prompts and responses into safety risk categories, including toxicity, hate speech, and hallucinations. Trained on a curated dataset, Llama Guard achieves performance on par with or exceeding existing moderation tools like OpenAI's Moderation API and ToxicChat. Its instruction-tuned architecture allows for customization, enabling developers to adapt its taxonomy and output formats to specific use cases. Llama Guard is part of Meta's broader "Purple Llama" initiative, which combines offensive and defensive security strategies to responsibly deploy generative AI models. The model weights are publicly available, encouraging further research and adaptation to meet evolving AI safety needs.

About

Shieldstral is a 3B open-weights, policy-adaptive multimodal safety classifier designed to evaluate text, images, and text-plus-image content using policies defined at inference time. Instead of relying on a fixed taxonomy of harm categories, it frames moderation as a binary question-answering task: users provide an instruction describing the evaluation context and strictness, a yes-or-no safety question, and the content to judge. The model reads the “yes” and “no” logits and converts them into a continuous, calibrated safety score, allowing applications to threshold or rank results by confidence rather than depend on a single discrete label. This formulation unifies prompt classification, response moderation, refusal detection, toxicity detection, and multimodal safety in one interface, while letting teams adapt policies without retraining the model. Shieldstral can evaluate prompts, responses, prompt-response pairs, images, and images with accompanying text.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Anyone searching for a tool to implement customizable safety measures in their generative AI applications

Audience

AI platform teams that need customizable, multimodal content moderation without retraining a separate safety model for every policy

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

No information available.
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Meta
Founded: 2004
United States
ai.meta.com/research/publications/llama-guard-llm-based-input-output-safeguard-for-human-ai-conversations/

Company Information

Mistral AI
Founded: 2023
France
mistral.ai/news/shieldstral/

Alternatives

Llama 3

Llama 3

Meta

Alternatives

Categories

Categories

Integrations

Llama
Mistral AI
Nebius Token Factory
OpenAI

Integrations

Llama
Mistral AI
Nebius Token Factory
OpenAI
Claim Llama Guard and update features and information
Claim Llama Guard and update features and information
Claim Shieldstral and update features and information
Claim Shieldstral and update features and information