Mixtral Sparse Mixture of Experts Model
Mistral Small 3.1 | Mistral AI

Mistral Small 3.1 | Mistral AI

3/17/2025

What this post added

Mistral Small 3.1 is introduced, featuring improved text performance, multimodal understanding, and an expanded context window of up to 128k tokens. It demonstrates performance exceeding comparable models like Gemma 3 and GPT-4o Mini, with inference speeds of 150 tokens per second. The post details instruct and pretrained performance across text, multimodal, multilingual, and long context benchmarks. It highlights use cases such as conversational assistance, image understanding, and function calling, emphasizing its suitability for on-device deployment and fine-tuning for specialized domains. The release includes both base and instruct checkpoints.

Read the original post ↗