MiniMax: MiniMax-H3
MiniMax
Jul 29, 2026
minimax/minimax-h3
Context Length7,000
Input Price-
Output Price$0.08/video
Max Output
Blended Price-

MiniMax-H3 is an omni-modal video generation model that produces 4-15 second clips at up to 2K resolution with native synchronized stereo audio in a single pass. It accepts a text prompt plus optional first-frame and last-frame images, and up to 9 reference images, 3 reference videos, and 3 reference audio clips, making it suited to storyboard continuation, style and subject transfer, and audio-driven performance.

Input Modalitiestext, image, video, audio
Output Modalitiesvideo
Modalitytext+image+video+audio->video
TokenizerMiniMaxTokenizer
Instruct Typenone
Provider Details

minimax

Input Cost-
Output Cost$0.08/video
Context Length7,000
Max Output4,096
Video Generation Pricing
Length768P2KDefault
1$0.080$0.130$0.080
Supported Parameters
Parameter
Type
Minimax
Media (Generation Controls)
aspectRatioOutput image aspect ratio.
string
"adaptive", "21:9", "16:9", "4:3", "1:1", "3:4", "9:16"
Other
callback_url
string
-
first_frame
string
-
image
string
-
images
array
max 9 items
last_frame
string
-
length
number
4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15
reference_audios
array
max 3 items
reference_videos
array
max 3 items
resolution
string
"768P", "2K"
Make Your First API Call

Get started by making your first request to FastRouter. Use your favorite SDK or make direct API calls.

Copy

Before you run this code:

  1. Replace <FASTROUTER_API_KEY> with your actual API key
  2. Run the code, then check your dashboard for analytics
Frequently Asked Questions

More models from MiniMax
minimax/minimax-m2.5

MiniMaxAI/MiniMax-M2.5 is a 229B-parameter Mixture-of-Experts (MoE) language model with 10B active parameters per token, designed as the world's first production-level model natively optimized for agentic scenarios like coding, tool use, search, and office productivity.

Model scores sourced from ArtificialAnalysis