MiniMax

MiniMax M3

Built for long-horizon coding, tool use, and multi-turn production collaboration, extending beyond MiniMax M2.7 with multimodal inputs while keeping a 1M-token context window. Compared with Claude, Gemini, and GLM-5.2 long-context agent models, it is better for putting text, image, and video materials into one workflow.

  • Reasoning
  • Tool use
  • Function calling
  • Structured output
  • Long context

AI Chat

Send a prompt and see the response here.

MiniMax-M3
Current conversation
How can I help you?

Start a conversation with this model and ask follow-up questions.

Uses your account balance at the model’s current rates.

Insufficient balance

Your account balance is insufficient for this generation. Top up, then return here to try again.

USD /1M tokens

Pricing

GroupTierInputOutputCache Read
default≤512K$0.315$1.26$0.063
default>1M$0.63$2.52$0.126
default
≤512K
Input$0.315Output$1.26Cache Read$0.063
>1M
Input$0.63Output$2.52Cache Read$0.126

Public list prices are shown in USD. Final charges may vary by account group and usage tier.

API

Code examples

curl --location --request POST 'https://api.tokenhot.ai/v1/chat/completions' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data-raw '{
    "model": "MiniMax-M3",
    "messages": [
        {
            "role": "user",
            "content": "What model are you"
        }
    ]
}'

FAQ

Frequently asked questions

What is MiniMax M3 best suited for?

MiniMax M3 is suited to the input, output, and capability types listed on this page. Test production workloads before choosing it for a critical system.

How is MiniMax M3 priced?

Pricing is shown in the pricing section above. Actual cost depends on usage volume, account group, and the request payload.

How can I call MiniMax M3?

Use the model ID shown above with one of the supported API protocols. Code examples are provided when usage data is available.

Get started

Build with MiniMax M3

Use in Console