icon

xiaomi-mimo-v2-pro-free

New
xiaomi-mimo-v2-pro-free is the open free version of xiaomi-mimo-v2-pro. To maintain reliable service, each account is limited to 5 requests per minute, 500 requests per day, and 1 million tokens per day.
Pricing Details
  • Input Tokens: $0.00 /M tokens
  • Output Tokens: $0.00 /M tokens
  • Cache Tokens: $0.00 /M tokens
Input Modalities
  • Text
  • Vision
  • Audio
  • Video
Output Modalities
  • Text
Features
  • Web
Tags
  • Free
API Usage Examples
Python
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AIHUBMIX_API_KEY"],
    base_url="https://aihubmix.com/v1",
)

response = client.chat.completions.create(
    model="xiaomi-mimo-v2-pro-free",
    messages=[
      {
        "role": "user",
        "content": "Hello, how are you?"
      }
    ],
    max_tokens=1024,
    stream=False,
)

print(response.choices[0].message.content)
FAQ
What is xiaomi-mimo-v2-pro-free?
xiaomi-mimo-v2-pro-free is the open free version of xiaomi-mimo-v2-pro. To maintain reliable service, each account is limited to 5 requests per minute, 500 requests per day, and 1 million tokens per day.
More from Xiaomi
  • Input: $ 0.155 /M Tokens
  • Output: $ 0.31 /M Tokens
  • Web Search: $0.005/request

MiMo-V2.5 is a native, fully multimodal large model designed for agent scenarios; it can see, hear, and read, and translate understanding into action. It has over 1 trillion total parameters (42B active parameters), employs an innovative hybrid-attention architecture, and supports an ultra-long 1M context length. Built on a powerful model base, we continuously scale compute across broader agent scenarios, further expanding the agent’s action space and achieving an important generalization from coding to claw.

Input:$ 0.48 /M Tokens
Output:$ 0.96 /M Tokens
Context:1M
Latency:1.673 S
Throughput:41 TPS

MiMo-V2.5-Pro is Xiaomi's most powerful model to date. In areas such as general agent capabilities, complex software engineering, and long-horizon tasks, it can now directly compete with the world's top agent models (Claude Opus 4.6, GPT-5.4). Compared with the previous-generation MiMo-V2-Pro, it achieves an all-around leap forward.

Input:$ 0.08 /M Tokens
Output:$ 0.16 /M Tokens
Context:-
Latency:-
Throughput:-

Only supports OpenAI-compatible formats.

Input:$ 0.2 /M Tokens
Output:$ 0.4 /M Tokens
Context:-
Latency:-
Throughput:-

Only supports OpenAI-compatible formats.

Input:$ 0.08 /M Tokens
Output:$ 0.4 /M Tokens
Context:-
Latency:-
Throughput:-

Only supports OpenAI-compatible formats.

Input:$ 0.2 /M Tokens
Output:$ 0.6 /M Tokens
Context:-
Latency:-
Throughput:-

Only supports OpenAI-compatible formats.