MiniMax-M3 is a versatile multimodal foundation model developed by MiniMax that supports text, image, and video inputs to generate text outputs. With a massive context window of 1,048,576 tokens, it is capable of processing and understanding vast amounts of information. This model is highly optimized for complex tasks, making it exceptionally well-suited for coding and long-horizon agentic workflows.
Pricing
- Input Tokens: $0.000 /M tokens
- Output Tokens: $0.000 /M tokens
- Cache Read: $0.000 /M tokens
Input Modalities
- Text
- Vision
Output Modalities
- Text
Capabilities
- Long context
Try this model
Python
