Minimax logo

Minimax AI models

6 Minimax models on FastRouter, all behind one OpenAI-compatible API. Compare pricing, context windows and benchmarks, then open any model for its providers and code samples.

Filter in catalog
Models
6
Largest context
1M
MiniMax M3
Lowest input /1M
$0.30
MiniMax M3
Top intelligence
22.8
MiniMax M2.5

All Minimax models

MiniMax logo

MiniMax-H3 is an omni-modal video generation model that produces 4-15 second clips at up to 2K resolution with native synchronized stereo audio in a single pass. It accepts a text prompt plus optional first-frame and last-frame images, and up to 9 reference images, 3 reference videos, and 3 reference audio clips, making it suited to storyboard continuation, style and subject transfer, and audio-driven performance.

minimax/minimax-h3Jul 29, 2026
Context
7K
Price /1M
— in$0.08/vid out
Intel
—
Minimax logo

MiniMax M3 is a multimodal foundation model from MiniMax, built for coding, agentic workflows, long-context reasoning, and multimodal inputs. It supports up to 1 million tokens of context through MiniMax’s Sparse Attention (MSA) architecture and is positioned as a frontier model for long-horizon tasks.

minimax/minimax-m3May 31, 20263.1s latency
Context
1M
Price /1M
$0.30 in$1.20 out
Intel
—
Minimax logo

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity with self-improvement capabilities. Created on March 18, 2026, it integrates advanced agentic capabilities through multi-agent collaboration, enabling it to plan, execute, and refine complex tasks across dynamic environments.

minimax/minimax-m2.7Mar 18, 20264.7s latency
Context
205K
Price /1M
$0.30 in$1.20 out
Intel
22.8
Minimax logo

MiniMax-M2.7-HighSpeed is the ultra-fast variant of MiniMax's M2.7 flagship model, delivering approximately 100 tokens per second—3x faster than competitors like Claude Opus 4.6 (~33 tps) and GPT-5 (~40 tps)—while maintaining identical performance on complex tasks at a fraction of the cost.

minimax/minimax-m2.7-highspeedMar 18, 20265.7s latency
Context
205K
Price /1M
$0.60 in$2.40 out
Intel
—
MiniMax logo

MiniMaxAI/MiniMax-M2.5 is a 229B-parameter Mixture-of-Experts (MoE) language model with 10B active parameters per token, designed as the world's first production-level model natively optimized for agentic scenarios like coding, tool use, search, and office productivity.

minimax/minimax-m2.5Feb 12, 20264.2s latency
Context
197K
Price /1M
$0.30 in$1.20 out
Intel
22.8
Minimax logo

MiniMax-M2.5-HighSpeed (also called the Lightning variant) is the high-throughput version of MiniMax's 229B-parameter Mixture-of-Experts (MoE) model, delivering 100 tokens/second natively—roughly 2x faster than other frontier models—while maintaining identical capabilities to the standard 50 tokens/second version.

minimax/minimax-m2.5-highspeedFeb 12, 20265.8s latency
Context
205K
Price /1M
$0.60 in$2.40 out
Intel
—

Frequently asked questions

Model scores sourced from ArtificialAnalysis

Minimax AI Models: API Pricing & Benchmarks | FastRouter.ai