MiniMax H3 (minimax-h3) is a general-purpose omni-modal generation model developed by the Chinese AI company MiniMax. It can jointly understand text, images, video, and audio while generating videos of up to 15 seconds in 2K resolution with native stereo sound. It is designed for advertising, branding, e-commerce, product design, UI/UX, and gaming. Compared with Hailuo 01 and Hailuo 02, H3 evolves from specialized video generation into a unified model for multimodal creation, reference-based editing, text rendering, and motion transfer.
Pricing
Video generation price: 480p video $0.05112/s. 768P video $0.07744/s. For images, after more than five, each additional image $0.0282.
Input Modalities
- Text
- Vision
Output Modalities
- Video
Providers
Minimax mm-minimax-h3-max
Context0
Max output0
Latency-
Throughput-
Uptime
0.00% uptime 2 days ago
0.00% uptime yesterday
0.00% uptime today
Performance for minimax-h3-max
Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).
Uptime
Loading...
Latency
Loading...
Throughput
Loading...
