12/2/2025
What this post added
This post announces Mistral 3, a new generation of models. It details Mistral Large 3, a sparse mixture-of-experts model, and the Ministral 3 series. It highlights the collaboration with NVIDIA, vLLM, and Red Hat for optimized training and inference, including specific technologies like llm-compressor, TensorRT-LLM, SGLang, and support for disaggregated serving and speculative decoding. The post also mentions the multimodal and multilingual capabilities of the new models and their availability on various platforms.