MiniMax

M2.5 API

A cost-efficient model for real productivity agents, strongest in coding, tool use, search, and office-style deliverables rather than casual chat alone. It fits agents that plan and execute cross-file changes, research retrieval, and document/spreadsheet/presentation workflows; versus more expensive flagships, the main tradeoff is lower cost for near-frontier execution efficiency.

  • Reasoning
  • Tool use
  • Function calling
  • Structured output
  • Long context
Use in ConsoleMiniMax-M2.5

USD

Pricing

default
Standard Rate
Input$0.302/1M tokensCompletion Price$1.207698/1M tokensCache Read$0.029415/1M tokensCache Write$0.367504/1M tokens

Public list prices are shown in USD. Final charges may vary by account group and usage tier.

API

Code examples

curl --location --request POST 'https://api.tokenhot.ai/v1/chat/completions' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data-raw '{
    "model": "MiniMax-M2.5",
    "messages": [
        {
            "role": "user",
            "content": "What model are you"
        }
    ]
}'

MiniMax

Related models

FAQ

Frequently asked questions

What is M2.5 best suited for?

M2.5 is suited to the input, output, and capability types listed on this page. Test production workloads before choosing it for a critical system.

How is M2.5 priced?

Pricing is shown in the pricing section above. Actual cost depends on usage volume, account group, and the request payload.

How can I call M2.5?

Use the model ID shown above with one of the supported API protocols. Code examples are provided when usage data is available.

Get started

Build with M2.5

Use in Console
TokenHot

The frontier intelligence gateway. One API. 127 models. 0.2s latency. Pay only for what you use.

All systems normal · 99.997% uptime

Company

© 2026 TokenHot Inc. — Built for builders.