BlogsMistral AIMinistral Edge Models

Ministral Edge Models

Ministral Edge Models

1
posts
2024

Mistral AI introduces Ministral 3B and Ministral 8B, new state-of-the-art models optimized for on-device and edge computing. These models offer enhanced knowledge, commonsense reasoning, function-calling, and efficiency in the sub-10B parameter category. They support up to 128k context length, with Ministral 8B featuring an interleaved sliding-window attention pattern for improved inference speed and memory efficiency. Use cases include privacy-first local inference for applications like smart assistants, local analytics, and autonomous robotics, as well as acting as efficient intermediaries for function-calling in multi-step agentic workflows. Benchmarks show these models consistently outperform peers in their size category, including Gemma 2 and Llama 3 variants, and even surpass Mistral 7B on many metrics. The models are available via API on la Plateforme with competitive pricing and also offered under Mistral Commercial and Research Licenses for self-deployment, with support for lossless quantization.

2024

Un Ministral, des Ministraux | Mistral AI

10/16/2024

Introduction of Ministral 3B and Ministral 8B models, specifically designed for edge and on-device inference. Key technical features include support for 128k context length and an interleaved sliding-window attention pattern in Ministral 8B for efficient inference. The post presents benchmark results demonstrating superior performance compared to existing models in the sub-10B parameter class.