Codestral Mamba

Codestral Mamba

Mistral AI
Qwen3.8-27B

Qwen3.8-27B

Alibaba
+
+

Related Products

  • Google AI Studio
    30 Ratings
    Visit Website
  • Gemini Enterprise Agent Platform
    999 Ratings
    Visit Website
  • JetBrains Junie
    12 Ratings
    Visit Website
  • Retool
    593 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Careerminds
    46 Ratings
    Visit Website
  • Runpod
    230 Ratings
    Visit Website
  • LTX
    182 Ratings
    Visit Website
  • LendingPad
    302 Ratings
    Visit Website
  • RealEstateAPI (REAPI)
    54 Ratings
    Visit Website

About

As a tribute to Cleopatra, whose glorious destiny ended in tragic snake circumstances, we are proud to release Codestral Mamba, a Mamba2 language model specialized in code generation, available under an Apache 2.0 license. Codestral Mamba is another step in our effort to study and provide new architectures. It is available for free use, modification, and distribution, and we hope it will open new perspectives in architecture research. Mamba models offer the advantage of linear time inference and the theoretical ability to model sequences of infinite length. It allows users to engage with the model extensively with quick responses, irrespective of the input length. This efficiency is especially relevant for code productivity use cases, this is why we trained this model with advanced code and reasoning capabilities, enabling it to perform on par with SOTA transformer-based models.

About

Qwen3.8-27B is a compact open-weights model in Alibaba’s Qwen3.8 family, aimed at developers and researchers who want strong local AI performance without using the full Max-scale model. Reports from Alibaba’s Qwen3.8 launch state that Qwen3.8-27B was planned for open-weight release alongside Qwen3.8-Max, expanding access for builders working on AI applications. The model is positioned for coding, research, professional workflows, and local deployment scenarios where a 27B model can be more practical than frontier-scale systems. Qwen3.8’s broader launch emphasizes software development, document processing, data analysis, and professional “cowork” use cases. Qwen3.8-27B is especially relevant for teams that need a capable open model for experimentation, coding agents, assistant workflows, and self-hosted inference. Built for practical deployment, Qwen3.8-27B gives developers a smaller Qwen3.8 option for building AI tools, testing agents, and running advanced language model workflows.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Developers and anyone in search of an AI language model to generate code

Audience

Software developers, AI researchers, coding agent builders, local LLM users, startup teams, enterprise AI teams, infrastructure teams, data teams, and organizations that need an open-weights 27B model for coding assistance, agent testing, self-hosted inference, document processing, data analysis, workflow automation, model benchmarking, private deployment, and Qwen-family experimentation

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Open source
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 5.0 / 5
ease 5.0 / 5

Pros & Cons from Real Users

Pros

  • A 27B model is big enough to be useful for serious coding, reasoning, writing, and local-agent workflows, but still small enough that developers can realistically experiment with quantized versions on enthusiast hardware. That is the sweet spot for a lot of builders. Not everyone wants a massive cloud-only model, and not every workflow needs a 2T+ parameter system. A strong 27B model can be great for private coding help, local RAG, repo exploration, prompt testing, and lightweight agents. I also like the open-weight/local angle. Community reports around Qwen3.8-27B are already focused on GGUFs, MLX builds, VRAM needs, and local performance, which is exactly the kind of ecosystem momentum that makes a model useful beyond a demo.

Cons

  • The main downside is clarity. I would want a stable official model card, confirmed architecture details, benchmarks, license info, and serving recommendations before treating it as a production-ready model.

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Mistral AI
France
mistral.ai/news/codestral-mamba/

Company Information

Alibaba
Founded: 1999
China
qwen.ai

Alternatives

Mistral Code

Mistral Code

Mistral AI

Alternatives

GLM-5.3

GLM-5.3

Z.ai
Qwen3.8-Max

Qwen3.8-Max

Alibaba
Falcon Mamba 7B

Falcon Mamba 7B

Technology Innovation Institute (TII)
Qwen3.6

Qwen3.6

Alibaba
Mistral NeMo

Mistral NeMo

Mistral AI
Qwen2

Qwen2

Alibaba

Categories

Categories

Integrations

Hugging Face
Alibaba Cloud
Alibaba Cloud Model Studio
Cherry Studio
Diaflow
Elixir
Graydient AI
JavaScript
Melies
Msty
Nutanix Enterprise AI
Overseer AI
PostgresML
Qwen
Qwen Code
QwenCloud
R
Ragas
Ruby
Verta

Integrations

Hugging Face
Alibaba Cloud
Alibaba Cloud Model Studio
Cherry Studio
Diaflow
Elixir
Graydient AI
JavaScript
Melies
Msty
Nutanix Enterprise AI
Overseer AI
PostgresML
Qwen
Qwen Code
QwenCloud
R
Ragas
Ruby
Verta
Claim Codestral Mamba and update features and information
Claim Codestral Mamba and update features and information
Claim Qwen3.8-27B and update features and information
Claim Qwen3.8-27B and update features and information