Ministral 8B

Ministral 8B

Mistral AI
Mu

Mu

Microsoft
+
+

Related Products

  • Google AI Studio
    30 Ratings
    Visit Website
  • Gemini Enterprise Agent Platform
    985 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Google Cloud Speech-to-Text
    366 Ratings
    Visit Website
  • Runpod
    220 Ratings
    Visit Website
  • RaimaDB
    12 Ratings
    Visit Website
  • ManageEngine EventLog Analyzer
    211 Ratings
    Visit Website
  • Time Management from ISGUS
    27 Ratings
    Visit Website
  • TimeControl
    1 Rating
    Visit Website
  • LendingPad
    302 Ratings
    Visit Website

About

Mistral AI has introduced two advanced models for on-device computing and edge applications, named "les Ministraux": Ministral 3B and Ministral 8B. These models excel in knowledge, commonsense reasoning, function-calling, and efficiency within the sub-10B parameter range. They support up to 128k context length and are designed for various applications, including on-device translation, offline smart assistants, local analytics, and autonomous robotics. Ministral 8B features an interleaved sliding-window attention pattern for faster and more memory-efficient inference. Both models can function as intermediaries in multi-step agentic workflows, handling tasks like input parsing, task routing, and API calls based on user intent with low latency and cost. Benchmark evaluations indicate that les Ministraux consistently outperforms comparable models across multiple tasks. As of October 16, 2024, both models are available, with Ministral 8B priced at $0.1 per million tokens.

About

Mu is a 330-million-parameter encoder–decoder language model designed to power the agent in Windows settings by mapping natural-language queries to Settings function calls, running fully on-device via NPUs at over 100 tokens per second while maintaining high accuracy. Drawing on Phi Silica optimizations, Mu’s encoder–decoder architecture reuses a fixed-length latent representation to cut computation and memory overhead, yielding 47 percent lower first-token latency and 4.7× higher decoding speed on Qualcomm Hexagon NPUs compared to similar decoder-only models. Hardware-aware tuning, including a 2/3–1/3 encoder–decoder parameter split, weight sharing between input and output embeddings, Dual LayerNorm, rotary positional embeddings, and grouped-query attention, enables fast inference at over 200 tokens per second on devices like Surface Laptop 7 and sub-500 ms response times for settings queries.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Anyone looking for a tool providing efficient, low-latency AI models to manage their agentic workflows

Audience

Developers seeking a solution to navigate and configure system settings through natural language

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Open source
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Mistral AI
Founded: 2023
France
mistral.ai/news/ministraux/

Company Information

Microsoft
Founded: 1975
United States
blogs.windows.com/windowsexperience/2025/06/23/introducing-mu-language-model-and-how-it-enabled-the-agent-in-windows-settings/

Alternatives

Ministral 3B

Ministral 3B

Mistral AI

Alternatives

CodeT5

CodeT5

Salesforce
GLM-OCR

GLM-OCR

Z.ai
LFM2

LFM2

Liquid AI
Mistral 7B

Mistral 7B

Mistral AI
Mistral NeMo

Mistral NeMo

Mistral AI
Falcon-7B

Falcon-7B

Technology Innovation Institute (TII)

Categories

Categories

Integrations

APIPark
Airtrain
Diaflow
GaiaNet
Graydient AI
Groq
Kiin
Mammouth AI
Microsoft Foundry Agent Service
Mirascope
OpenPipe
PostgresML
Prompt Security
PromptPal
Respan
Simplismart
SydeLabs
Unify AI
Wordware
bolt.diy

Integrations

APIPark
Airtrain
Diaflow
GaiaNet
Graydient AI
Groq
Kiin
Mammouth AI
Microsoft Foundry Agent Service
Mirascope
OpenPipe
PostgresML
Prompt Security
PromptPal
Respan
Simplismart
SydeLabs
Unify AI
Wordware
bolt.diy
Claim Ministral 8B and update features and information
Claim Ministral 8B and update features and information
Claim Mu and update features and information
Claim Mu and update features and information