Mistral Medium 3 is Mistral's flagship multimodal model supporting text and image inputs with a 131K token context window.
| Provider | Input / 1M | Output / 1M | Cached / 1M |
|---|---|---|---|
| $0.400 | $2.00 | $0.040 |
Prices updated daily. Last check: Sep 8, 2026
Input, output, and batch rates, plus alternatives, for one provider at a time.
Benchmarks measured Sep 2026. Scores are independent evaluations, not vendor-reported.
Mistral Medium 3 is designed for sophisticated applications requiring multimodal understanding and complex reasoning. Its combination of text and image processing makes it suitable for document analysis, research assistance, content analysis involving visual elements, and advanced chatbot applications. The 131K context window enables analysis of lengthy reports, academic papers, and extended conversations. Organizations requiring European-developed AI solutions may prefer this model for data sovereignty considerations. The flagship positioning makes it appropriate for demanding enterprise applications, research projects, and scenarios where high-quality reasoning is prioritized over tool integration.
Mistral Medium 3 pricing varies by provider and pricing type (standard vs batch). Check the pricing table above for current rates across all providers.
Mistral Medium 3 excels at complex reasoning tasks involving both text and images, such as document analysis, research assistance, and multimodal content understanding. Its 131K context window makes it particularly suitable for processing lengthy documents and maintaining extended conversations.
No, Mistral Medium 3 does not include tool calling capabilities. It focuses on direct text generation and multimodal understanding rather than external tool integration. For applications requiring function calling, consider other models in the Mistral family or alternative providers.