Ministral 8B is Mistral's lightweight multimodal model with 8 billion parameters, supporting text and image inputs with a 262K token context window.
| Provider | Input / 1M | Output / 1M | Cached / 1M |
|---|---|---|---|
| $0.150 | $0.150 | - | |
| $0.150 | $0.150 | $0.015 |
Prices updated daily. Last check: Sep 6, 2026
Input, output, and batch rates, plus alternatives, for one provider at a time.
Ministral 8B is well-suited for multimodal applications where efficiency and cost-effectiveness are priorities. Its combination of vision capabilities and extended context makes it ideal for document analysis workflows, content moderation systems that need to process both text and images, educational content review, and customer service applications involving visual materials. The model's lightweight nature makes it appropriate for high-volume deployments where multimodal understanding is needed but the computational requirements of larger models would be prohibitive. Organizations processing mixed media content, performing visual document analysis, or building applications that need to understand both textual and visual information will find this model's balance of capabilities and efficiency particularly valuable.
Ministral 8B pricing varies by provider and may differ for input versus output tokens. Check the pricing table above for current rates across all providers offering this model.
Ministral 8B excels at multimodal tasks requiring both text and image understanding, such as document analysis, content moderation, and visual question answering. Its 262K context window and lightweight architecture make it ideal for high-volume applications needing efficient multimodal processing.
No, Ministral 8B does not support tool calling or function calling capabilities. It focuses on text generation and multimodal understanding for direct response generation rather than external tool integration.