LLM API Pricing

MiniMax Models

Browse all 5 MiniMax models available on TokenHot.

One API key gives you access to 5 MiniMax models with live pricing and no separate provider account.

Minimax
MiniMax-M3
MiniMax
CONTEXT1M
Input$1.2600/M
Output$5.0400/M
Built for long-horizon coding, tool use, and multi-turn production collaboration, extending beyond MiniMax M2.7 with multimodal inputs while keeping a 1M-token context window. Compared with Claude, Gemini, and GLM-5.2 long-context agent models, it is better for putting text, image, and video materials into one workflow.
Input Type:
Output Type:
ReasoningTool UseFunction CallingStructured OutputLong Context
Minimax
MiniMax-M2.7-highspeed
MiniMax
CONTEXT200K
Input$0.6300/M
Output$2.5200/M
Built for text generation, reasoning, and tool use, with stronger 200K long-context and agentic workflow stability than earlier MiniMax M2.x versions. Compared with Chinese peers such as Kimi, GLM, and Qwen, it fits productivity, coding assistance, and long-document processing.
Input Type:
Output Type:
Reasoning
Minimax
MiniMax-M2.7
MiniMax
CONTEXT200K
Input$0.3150/M
Output$1.2600/M
Built for text generation, reasoning, and tool use, with stronger 200K long-context and agentic workflow stability than earlier MiniMax M2.x versions. Compared with Chinese peers such as Kimi, GLM, and Qwen, it fits productivity, coding assistance, and long-document processing.
Input Type:
Output Type:
Reasoning
Minimax
MiniMax-M2.5
MiniMax
CONTEXT205K
Input$0.3020/M
Output$1.2077/M
A cost-efficient model for real productivity agents, strongest in coding, tool use, search, and office-style deliverables rather than casual chat alone. It fits agents that plan and execute cross-file changes, research retrieval, and document/spreadsheet/presentation workflows; versus more expensive flagships, the main tradeoff is lower cost for near-frontier execution efficiency.
Input Type:
Output Type:
ReasoningTool UseFunction CallingStructured OutputLong Context
Minimax
MiniMax-H3
MiniMax
PER SEC$0.0800/s