LLM API Pricing

MiniMax Models

Browse all 5 MiniMax models available on TokenHot.

One API key gives you access to 5 MiniMax models with live pricing and no separate provider account.

Minimax
MiniMax
PER SEC$0.0800/s
An omni-modal video generation model that can combine text with image, video, and audio references to create video assets with synchronized audio. It is suited to short-form video, advertising, and reference-driven character or camera work; it is not a general chat model, so prioritize visual consistency, motion, and audio-visual quality.
Input Type:
Output Type:
Multimodal Output
Minimax
MiniMax
CONTEXT1M
Input$0.3150/M
Output$1.2600/M
Built for long-horizon coding, tool use, and multi-turn production collaboration, extending beyond MiniMax M2.7 with multimodal inputs while keeping a 1M-token context window. Compared with Claude, Gemini, and GLM-5.2 long-context agent models, it is better for putting text, image, and video materials into one workflow.
Input Type:
Output Type:
ReasoningTool UseFunction CallingStructured OutputLong Context
Try model
CONTEXT200K
Input$0.6300/M
Output$2.5200/M
Built for text generation, reasoning, and tool use, with stronger 200K long-context and agentic workflow stability than earlier MiniMax M2.x versions. Compared with Chinese peers such as Kimi, GLM, and Qwen, it fits productivity, coding assistance, and long-document processing.
Input Type:
Output Type:
Reasoning
Try model
Minimax
CONTEXT200K
Input$0.3150/M
Output$1.2600/M
Built for text generation, reasoning, and tool use, with stronger 200K long-context and agentic workflow stability than earlier MiniMax M2.x versions. Compared with Chinese peers such as Kimi, GLM, and Qwen, it fits productivity, coding assistance, and long-document processing.
Input Type:
Output Type:
Reasoning
Try model
Minimax
CONTEXT205K
Input$0.3020/M
Output$1.2077/M
A cost-efficient model for real productivity agents, strongest in coding, tool use, search, and office-style deliverables rather than casual chat alone. It fits agents that plan and execute cross-file changes, research retrieval, and document/spreadsheet/presentation workflows; versus more expensive flagships, the main tradeoff is lower cost for near-frontier execution efficiency.
Input Type:
Output Type:
ReasoningTool UseFunction CallingStructured OutputLong Context
Try model