Average Ratings 0 Ratings
Average Ratings 1 Rating
Description
In honor of Cleopatra, whose magnificent fate concluded amidst the tragic incident involving a snake, we are excited to introduce Codestral Mamba, a Mamba2 language model specifically designed for code generation and released under an Apache 2.0 license. Codestral Mamba represents a significant advancement in our ongoing initiative to explore and develop innovative architectures. It is freely accessible for use, modification, and distribution, and we aspire for it to unlock new avenues in architectural research. The Mamba models are distinguished by their linear time inference capabilities and their theoretical potential to handle sequences of infinite length. This feature enables users to interact with the model effectively, providing rapid responses regardless of input size. Such efficiency is particularly advantageous for enhancing code productivity; therefore, we have equipped this model with sophisticated coding and reasoning skills, allowing it to perform competitively with state-of-the-art transformer-based models. As we continue to innovate, we believe Codestral Mamba will inspire further advancements in the coding community.
Description
Qwen3.8-27B is a 27B-class open-weights model associated with Alibaba’s Qwen3.8 model family. Alibaba’s Qwen3.8 release positioned the broader family as a top-tier large language model system optimized for coding and professional cowork scenarios. Reports indicate that Qwen3.8-27B was planned to be released as open weights alongside Qwen3.8-Max, giving developers and researchers a more accessible option than the full Max-scale model. The model is designed for users who want strong AI capability in a smaller, more deployable package. Qwen3.8-27B can support workflows such as coding assistance, AI agents, research tasks, document analysis, data work, and self-hosted experimentation. The larger Qwen3.8-Max release is described as targeting coding, research, professional work, and multimodal tasks, and Qwen3.8-27B appears to serve builders who need a more practical model size for local or private infrastructure. QwenCloud documentation confirms that the Qwen3.8 generation includes modern capabilities such as thinking, function calling, built-in tools, and structured output for the Max model. Community discussion and third-party coverage also highlight interest in running Qwen3.8-27B through GGUF and local inference workflows. By combining open-weight accessibility, a 27B-class footprint, Qwen3.8-era capability, and developer-focused use cases, Qwen3.8-27B gives teams a practical model for coding and agentic experimentation.
API Access
Has API
Yes
API Access
Has API
Yes
Integrations
Hugging Face
Yes
AnythingLLM
Yes
Arize Phoenix
Yes
CSS
Yes
Diaflow
Yes
F#
Yes
Fleak
Yes
GaiaNet
Yes
Go
Yes
Hermes Agent
No
Integrations
Hugging Face
Yes
AnythingLLM
No
Arize Phoenix
No
CSS
No
Diaflow
No
F#
No
Fleak
No
GaiaNet
No
Go
No
Hermes Agent
Yes
Pricing Details
Free
Open source
Free Trial
No
Free Version
Yes
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Vendor Details
Company Name
Mistral AI
Country
France
Website
mistral.ai/news/codestral-mamba/
Vendor Details
Company Name
Alibaba
Founded
1999
Country
China
Website
qwen.ai