Blogs›Mistral AI›Ministral Edge Models
Mistral AI introduces Ministral 3B and Ministral 8B, new state-of-the-art models optimized for on-device and edge computing. These models offer enhanced knowledge, commonsense reasoning, function-calling, and efficiency in the sub-10B parameter category. They support up to 128k context length, with Ministral 8B featuring an interleaved sliding-window attention pattern for improved inference speed and memory efficiency. Use cases include privacy-first local inference for applications like smart assistants, local analytics, and autonomous robotics, as well as acting as efficient intermediaries for function-calling in multi-step agentic workflows. Benchmarks show these models consistently outperform peers in their size category, including Gemma 2 and Llama 3 variants, and even surpass Mistral 7B on many metrics. The models are available via API on la Plateforme with competitive pricing and also offered under Mistral Commercial and Research Licenses for self-deployment, with support for lossless quantization.