Model catalog

MiniMax Models

Browse the AI models associated with MiniMax. Compare current and historical models by family, capabilities, context window, availability and intended use.

27 models tracked
27 Total models
17 Model families
5 Model types
26 Current / accessible
All models

MiniMax model catalog

Short text-to-video and image-to-video clips, cinematic experiments, advertising concepts, social content, and scenes with complex motion.

Type Multimodal
Reasoning 1/10
Speed 5/10
Multimodal Image input Media output
Status

Legacy or superseded; no longer prominent in MiniMax's current first-party model catalog

Output Historical launch configurations were priced per generated video rather than by tokens; documented configurations included 512p/6s, 512p/10s, 768p/6s, 768p/10s, and 1080p/6s.
View model →

Short cinematic videos, image animation, realistic human motion, stylized scenes, visual effects, advertising concepts, and social-media content.

Type Video Generation
Reasoning 2/10
Speed 6/10
Multimodal Image input Media output
Status

Legacy or superseded video model; still referenced in MiniMax consumer and platform offerings, while MiniMax H3 is the current primary video model in developer documentation.

Output $0.28 per 768P/6s clip; $0.56 per 768P/10s clip; $0.49 per 1080P/6s clip
View model →

Fast image-to-video generation, high-volume short-form content, social media clips, advertisements, and rapid creative iteration

Type Other
Reasoning 1/10
Speed 9/10
Multimodal Image input Media output
Status

Legacy; current availability should be verified

Input $0.19 per 768p 6-second clip; $0.32 per 768p 10-second clip; $0.33 per 1080p 6-second clip
View model →
MiniMax logo
Hailuo Video

MiniMax T2V-01

Short text-driven video concepts, storyboards, and early Hailuo-style cinematic experiments

Type Other
Reasoning 2/10
Speed 6/10
Media output
Status

Legacy and superseded; no longer listed in MiniMax's current primary video-generation catalog

View model →
MiniMax logo
Image-01

image-01

Prompt-based image generation, reference-guided variations, commercial visuals, and batch creative production

Type Other
Speed 7/10
Multimodal Image input Media output
Status

Current and accessible through MiniMax's image-generation console and API platform

View model →
MiniMax logo
MiniMax ASR

ASR 1.0

Multilingual audio transcription, meeting transcription, speaker-labeled transcripts, live captions, call analysis, and subtitle generation

Type Other
Reasoning 1/10
Speed 7/10
Audio input Streaming
Status

Current public model

Input $0.38 per hour of processed audio
Output No separate output charge; billing is based on input audio duration
View model →
MiniMax logo
MiniMax H3

H3 Max

Fast short-form text-to-video, image-to-video, reference-guided generation, and synchronized-audio production

Type Multimodal
Reasoning 2/10
Speed 10/10
Multimodal Image input Audio input
Status

Current; commercially hosted by fal

Output $0.05 per second at 480p; $0.08 per second at 768p; $0.16 per second at 1080p
View model →
MiniMax logo
MiniMax H3

MiniMax H3

Multimodal commercial video generation, reference-based editing, product and advertising content, short cinematic clips, and locally deployed 768p workflows

Type Multimodal
Reasoning 2/10
Speed 5/10
Multimodal Image input Audio input
Status

Current; open-weight release and hosted API available

Input $0.08 per second for 768p video output; $0.13 per second for 2K video output; H3-Context-IR: $0.90 per million input tokens
Output $0.08 per second for 768p video output; $0.13 per second for 2K video output; H3-Regenerate-2K: $0.05 per second of regenerated output; H3-Context-IR: $3.60 per million output tokens
View model →
MiniMax logo
MiniMax M2

MiniMax M2

Coding agents, multi-step tool workflows, long-context codebase analysis, research automation, and self-hosted experimentation

Type Coding
Context 197K
Reasoning 8/10
Speed 8/10
Tool use Streaming
Status

Legacy/open-weight model; current hosted availability and pricing are not confirmed in MiniMax's latest public model catalog

Input $0.30 per 1 million input tokens at launch; current pricing unverified
Output $1.20 per 1 million output tokens at launch; current pricing unverified
View model →
MiniMax logo
MiniMax M2

MiniMax M2.1

Multilingual software engineering, coding agents, tool-using workflows, application development, and cost-sensitive automation

Type Coding
Context 205K
Reasoning 8/10
Speed 8/10
Tool use Streaming
Status

Legacy but currently available

Input $0.30 per 1 million tokens; prompt-cache read $0.03 per 1 million tokens; prompt-cache write $0.375 per 1 million tokens
Output $1.20 per 1 million tokens
View model →
MiniMax logo
MiniMax M2

MiniMax M2.5

Coding agents, software engineering, search and browser agents, tool-calling workflows, office automation, and long-context technical work

Type Reasoning
Context 205K
Reasoning 9/10
Speed 9/10
Tool use Web search Fine-tuning
Status

Current and accessible; open-weight model; also available through the MiniMax API and MiniMax Agent

Input $0.15 per 1 million input tokens for standard-speed M2.5 at launch; current prices may vary by platform and endpoint
Output $1.20 per 1 million output tokens for standard-speed M2.5
View model →
MiniMax logo
MiniMax M2

MiniMax M2.7

Agentic software engineering, repository-level coding, production debugging, complex tool workflows, office document automation, and long-context professional tasks

Type Coding
Context 197K
Reasoning 8/10
Speed 8/10
Tool use
Status

Current; available through MiniMax API, MiniMax Agent, and downloadable open weights

Input $0.30 per 1 million tokens; cache read $0.06 per 1 million tokens; cache write $0.375 per 1 million tokens
Output $1.20 per 1 million tokens
View model →

Low-latency coding assistants, multilingual software development, tool-using agents, long-horizon workflows, and interactive office automation

Type Coding
Context 205K
Reasoning 8/10
Speed 9/10
Tool use Streaming
Status

Current and available through the MiniMax Open Platform API; historical model variant

Input ¥4.20 per 1M input tokens
Output ¥16.80 per 1M output tokens
View model →

Low-latency coding assistants, software-engineering agents, tool-using workflows, search tasks, and long-context productivity automation

Type Coding
Context 205K
Reasoning 8/10
Speed 10/10
Tool use Web search Fine-tuning
Status

Current and available; highspeed variant of MiniMax M2.5

Input $0.30 per 1 million tokens
Output $2.40 per 1 million tokens
View model →

Low-latency coding assistants, software-engineering agents, tool-calling workflows, and interactive developer applications

Type Coding
Context 205K
Reasoning 8/10
Speed 10/10
Tool use Streaming
Status

Current and available through the MiniMax API Platform; highspeed variant of MiniMax M2.7

Input $0.60 per million tokens; cache read $0.06 per million tokens; cache write $0.375 per million tokens
Output $2.40 per million tokens
View model →
MiniMax logo
MiniMax M3

MiniMax M3

Long-context coding agents, autonomous tool-using workflows, multimodal document and video analysis, and private deployment

Type Multimodal
Context 1M
Reasoning 9/10
Speed 8/10
Multimodal Image input Video input
Status

Active; open-weight and available through the MiniMax API

Input $0.30 per million tokens for context up to 512K; $0.60 per million tokens for context from 512K to 1M. Listed rates are subject to MiniMax's current pricing terms.
Output $1.20 per million tokens for context up to 512K; $2.40 per million tokens for context from 512K to 1M. Listed rates are subject to MiniMax's current pricing terms.
View model →
MiniMax logo
MiniMax Music

MiniMax Music 2.0

Generating complete songs with expressive vocals, lyrics, melodies, instrumental arrangements, duets, a cappella passages, and cinematic musical soundscapes.

Type Other
Reasoning 1/10
Speed 6/10
Media output
Status

Legacy or limited availability; MiniMax states that music models were no longer available through Token Plan from 2026-08-20, but a full model retirement date was not verified.

View model →
MiniMax logo
MiniMax Music

MiniMax Music 2.6

Text-to-music, instrumental generation, game and video scoring, detailed musical direction, and genre reinterpretation with Cover mode

Type Other
Speed 7/10
Multimodal Audio input Media output
Status

Legacy or superseded; current public MiniMax navigation highlights Music 3.0, and Music 2.6 availability should be verified before use

View model →
MiniMax logo
MiniMax Music

MiniMax Music 3.0

Local generation of complete songs from lyrics and structured musical descriptions

Type Other
Context 5K
Reasoning 1/10
Speed 4/10
Media output
Status

Open-weight model; MiniMax Token Plan access discontinued on 2026-08-20

View model →
MiniMax logo
MiniMax Music

MiniMax Music Cover

Reinterpreting existing songs in new genres, vocal styles, arrangements, and production directions while preserving the source melody

Type Other
Reasoning 1/10
Speed 6/10
Multimodal Audio input Media output
Status

Current; specialized music-cover model introduced with MiniMax Music 2.6

Input Not applicable to token pricing; audio-generation pricing varies by MiniMax platform or deployment
Output Not applicable to token pricing; charged per generated audio result where applicable
View model →
MiniMax logo
MiniMax Speech

Speech-02-HD

High-quality multilingual voiceovers, audiobooks, narration, digital characters, advertising, education, and zero-shot voice cloning

Type Other
Context 10K
Reasoning 1/10
Speed 6/10
Multimodal Audio input Media output
Status

Legacy or older generation; still accessible through some partner platforms, but not listed as a core model on MiniMax's current global pricing page

Input CNY 3.5 per 10,000 characters on Alibaba Cloud Model Studio in China; current direct MiniMax pricing for this exact model is not verified
View model →

Low-latency multilingual text-to-speech, streaming voice agents, interactive applications, expressive narration, and voice cloning

Type Other
Reasoning 1/10
Speed 9/10
Multimodal Audio input Media output
Status

Legacy or superseded; exact current first-party availability is unclear

View model →

Real-time voice agents, conversational assistants, customer-service automation, interactive characters, multilingual speech, and low-latency text-to-speech

Type Other
Reasoning 1/10
Speed 9/10
Media output Streaming
Status

Legacy or superseded; current first-party availability is unverified

View model →
MiniMax logo
Speech 2.6

Speech-2.6-HD

High-quality voiceovers, audiobooks, narration, localization, e-learning, game dialogue, accessibility audio, and production speech

Type Other
Speed 7/10
Media output
Status

Legacy or transition-era model; not prominently listed in MiniMax's current first-party speech catalog as of September 25, 2026

View model →
MiniMax logo
Speech 2.8

Speech-2.8-HD

High-quality expressive narration, audiobooks, podcasts, advertising, character voices, multilingual speech, and applications prioritizing audio fidelity over the lowest latency

Type Other
Speed 8/10
Media output Streaming
Status

Current and available through the MiniMax API

Input $100 per 1 million characters for text-to-audio usage
View model →

Real-time text-to-speech, voice assistants, conversational agents, interactive applications, multilingual narration, gaming characters and expressive voice experiences

Type Other
Reasoning 1/10
Speed 9/10
Media output Streaming
Status

Current and accessible through the MiniMax Open Platform API

Input $60 per 1 million characters
View model →

Text-to-video generation with explicit cinematic camera-movement direction, short advertising concepts, storyboards, and controlled visual experiments

Type Other
Reasoning 1/10
Speed 6/10
Media output
Status

Legacy or limited availability; current official catalog status and continued first-party access are not clearly verified

View model →