Amux
All models

MiniMax: MiniMax H3 Max

Active

MiniMax: MiniMax H3 Max is a video generation model jointly released by MiniMax and fal.ai. The model was further trained by fal.ai based on MiniMax H3 and optimized for high-speed generation. It supports mainstream 480P and 768P output, generating videos faster than MiniMax H3. Currently, it supports 'text-to-video' and 'image-to-video', and will later support 'reference image generation'.

Input:
Output:
Context length:
13K
Published:
2026-09-02
vision

Providers

Same model, different providers. Automatic routing picks the cheapest healthy one within the same quality tier.

Provider
minimax
480P$0.05/second768P$0.08/second
ReadWrite
13K

Availability

24 hours

Success rate of the requests Amux actually sent to each provider. Client errors are excluded from the denominator — a malformed request isn't the provider's fault.

Activity

24 hours

Token usage and cost for this model over time, split by the provider that served each request.

Split by the provider that served each request. Amounts match what is actually billed — calls that failed upstream are not charged, so they are excluded.

Related models

More models from MiniMax

MiniMax: MiniMax H313K context1 providerMiniMax: MiniMax H3 is a lightweight open source weighted video generation model developed by MiniMax. It is designed for precise multimodal editing and controlled content generation, including command-driven editing, text and brand rendering, and video-to-video motion transfer. The model is suitable for commercial creative workflows such as advertising, e-commerce, games and interface design, and supports native audiovisual output for reference-driven generation.
MiniMax: MinMax M31M contextfrom $0.3 / M tokens1 providerMiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding, and tool use. It is built on MiniMax Sparse Attention (MSA), which replaces full attention with KV-block selection to cut per-token compute at long context — roughly 1/20 the cost of the previous generation at 1M tokens, with substantially faster prefill and decode while retaining quality across most tasks.
MiniMax: MinMax M2.7205K contextfrom $0.21 / M tokens1 providerMiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent collaboration, enabling it to plan, execute, and refine complex tasks across dynamic environments.