Model catalog
One endpoint, every model. Compare context windows, capabilities and real per-token pricing side by side, then swap the model string in your request — nothing else changes.
- 351
- Models
- 35
- Providers
- 11
- Free to try
- $0.0001
- Cheapest input / 1M
9 of 351 models
- Language
MiniMax M3
minimax/minimax-m3
MiniMax-M3 is a frontier-class foundation model that unites the three capabilities defining today's frontier: a 1M-token context window, frontier coding and agentic performance, and native multimodality — the first open-weight model to deliver all three in a single system.
- Context
- 1M
- In / 1M
- $0.300
- Out / 1M
- $1.20
- Reasoning
- Tool use
- Vision
- +2
- Language
Minimax M2.7
minimax/minimax-m2.7
M2.7 delivers outstanding performance in real-world software engineering, including end-to-end full project delivery, log analysis and bug troubleshooting, code security, machine learning, and more.
- Context
- 205K
- In / 1M
- $0.300
- Out / 1M
- $1.20
- Reasoning
- Tool use
- Implicit caching
- Language
MiniMax M2.7 High Speed
minimax/minimax-m2.7-highspeed
M2.7 Highspeed: Same performance, faster and more agile (output speed approximately 100 tps)
- Context
- 205K
- In / 1M
- $0.600
- Out / 1M
- $2.40
- Reasoning
- Tool use
- Implicit caching
- Language
MiniMax M2.5
minimax/minimax-m2.5
MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. It is capable of handling the entire development process of various complex systems. It covers full-stack projects across multiple platforms including Web, Android, iOS, Windows, and Mac, encompassing server-side APIs, functional logic, and databases.
- Context
- 205K
- In / 1M
- $0.300
- Out / 1M
- $1.20
- Reasoning
- Tool use
- Implicit caching
- Language
MiniMax M2.5 High Speed
minimax/minimax-m2.5-highspeed
M2.5 highspeed: Same performance, faster and more agile (output speed approximately 100 tps)
- Context
- 205K
- In / 1M
- $0.600
- Out / 1M
- $2.40
- Reasoning
- Tool use
- Implicit caching
- Language
MiniMax M2.1
minimax/minimax-m2.1
MiniMax 2.1 is MiniMax's latest model, optimized specifically for robustness in coding, tool use, instruction following, and long-horizon planning.
- Context
- 205K
- In / 1M
- $0.300
- Out / 1M
- $1.20
- Reasoning
- Tool use
- Implicit caching
- Language
MiniMax M2.1 Lightning
minimax/minimax-m2.1-lightning
MiniMax-M2.1-lightning is a faster version of MiniMax-M2.1, offering the same performance but with significantly higher throughput (output speed ~100 TPS, MiniMax-M2 output speed ~60 TPS).
- Context
- 205K
- In / 1M
- $0.300
- Out / 1M
- $2.40
- Reasoning
- Tool use
- Implicit caching
- Video
MiniMax H3
minimax/minimax-h3
H3 is a next-generation open-weights, general-purpose multimodal video model. Rather than being limited to specialized tasks such as generating, editing, or referencing, H3 understands multimodal contexts that bring together text, images, video, and audio. This enables it to interpret creative intent in a unified way and deliver more natural, coherent generation and expression.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Vision
- Video generation
- Language
MiniMax M2
minimax/minimax-m2
MiniMax-M2 redefines efficiency for agents. It is a compact, fast, and cost-effective MoE model (230 billion total parameters with 10 billion active parameters) built for elite performance in coding and agentic tasks, all while maintaining powerful general intelligence.
- Context
- 205K
- In / 1M
- $0.300
- Out / 1M
- $1.20
- Reasoning
- Tool use
