BlogsMistral AIBatch API for Model Inference

Batch API for Model Inference

Batch API for Model Inference

1
posts
2024

Mistral AI introduces a Batch API for its models, offering a more cost-efficient way (50% lower cost) to process high-volume requests compared to synchronous API calls. This feature is ideal for applications prioritizing data volume over synchronous responses, enabling users to upload batch files and download processed outputs. Popular use cases include bulk customer feedback analysis, document summarization and translation, vector embedding generation, and data labeling. The Batch API is available for all models on La Plateforme and will be extended to cloud provider partners. Usage is limited to 1 million ongoing requests per workspace.

2024

Mistral batch API

11/7/2024

Introduces a new Batch API for Mistral models, enabling asynchronous, high-volume request processing at a reduced cost. This addresses the need for efficient processing of large datasets for tasks like sentiment analysis, bulk summarization/translation, and embedding generation, by allowing users to upload batch files and retrieve results asynchronously.