MiniMax 01 is MiniMax's flagship multimodal model with text and image capabilities, featuring a massive 1M+ token context window for processing extensive documents and conversations.
| Provider | Input / 1M | Output / 1M |
|---|---|---|
| $0.200 | $1.10 |
Prices updated daily. Last check: Sep 6, 2026
Input, output, and batch rates, plus alternatives, for one provider at a time.
MiniMax 01 is well-suited for applications requiring extensive context retention and multimodal processing. Its massive context window makes it ideal for long document analysis, legal document review, research tasks involving large datasets, and extended technical consultations. The vision capabilities enable document OCR and analysis, image-based question answering, and multimodal content creation workflows. Organizations needing to process lengthy transcripts, analyze large codebases, or maintain context across very long conversations would benefit from this model's context capacity. However, applications requiring tool integration or function calling would need to implement those capabilities externally.
MiniMax 01 pricing varies by provider and usage patterns. Check the pricing table above for current rates across all available providers offering this model.
MiniMax 01 excels at tasks requiring extensive context retention and multimodal processing, such as long document analysis, research across large datasets, extended technical conversations, and vision-language tasks involving both text and images.
No, MiniMax 01 does not have built-in tool calling capabilities. Applications requiring function calling or API integration would need to implement these features through external orchestration layers.