
xiaomi-mimo-v2-pro-free
- Input Tokens: $0.00 /M tokens
- Output Tokens: $0.00 /M tokens
- Cache Tokens: $0.00 /M tokens
- Text
- Vision
- Audio
- Video
- Text
- Web
- Free

- Input: $ 0.155 /M Tokens
- Output: $ 0.31 /M Tokens
- Web Search: $0.005/request
MiMo-V2.5 is a native, fully multimodal large model designed for agent scenarios; it can see, hear, and read, and translate understanding into action. It has over 1 trillion total parameters (42B active parameters), employs an innovative hybrid-attention architecture, and supports an ultra-long 1M context length. Built on a powerful model base, we continuously scale compute across broader agent scenarios, further expanding the agent’s action space and achieving an important generalization from coding to claw.

MiMo-V2.5-Pro is Xiaomi's most powerful model to date. In areas such as general agent capabilities, complex software engineering, and long-horizon tasks, it can now directly compete with the world's top agent models (Claude Opus 4.6, GPT-5.4). Compared with the previous-generation MiMo-V2-Pro, it achieves an all-around leap forward.

Only supports OpenAI-compatible formats.

Only supports OpenAI-compatible formats.

Only supports OpenAI-compatible formats.

Only supports OpenAI-compatible formats.
