7/15/2025
What this post added
This post introduces the Voxtral models, detailing their architecture (24B and 3B variants), performance benchmarks against leading models (Whisper, GPT-4o mini, Gemini 2.5 Flash, ElevenLabs Scribe), and key features like long-form context, multilingual support, and function calling. It highlights the technical trade-offs addressed by Voxtral in the speech intelligence market and outlines enterprise-grade features and future development plans for audio capabilities.