Xiaomi MiMo

Mimo V2.5 API

Built for agents, coding, and long-context tasks, offering more complete multimodal understanding, a million-token context window, and long-horizon reasoning than MiMo V2. Compared with Chinese peers such as Qwen, GLM, and DeepSeek, its strengths are Xiaomi ecosystem fit and engineering automation scenarios.

  • Reasoning
  • Tool use
  • Structured output
  • Long context

USD

Pricing

default
imput<256K
Input$0.47/1M tokensCompletion Price$2.36/1M tokensCache Read$0.088/1M tokens
imput>=256K
Input$0.94/1M tokensCompletion Price$4.7/1M tokensCache Read$0.176/1M tokens

Public list prices are shown in USD. Final charges may vary by account group and usage tier.

Xiaomi MiMo

Related models

XiaomiMiMoXiaomi MiMoMimo V2.5 PRO

Built for agents, coding, and long-context tasks, offering more complete multimodal understanding, a million-token context window, and long-horizon reasoning than MiMo V2. Compared with Chinese peers such as Qwen, GLM, and DeepSeek, its strengths are Xiaomi ecosystem fit and engineering automation scenarios.

$1.17/1M tokens
XiaomiMiMoXiaomi MiMoMimo V2.5 TTS Voicedesign

Built for one-sentence voice design and text-to-speech, with stronger V2.5-series natural-language voice creation, fine-grained pace/emotion/tone control, and low-friction voice production than MiMo V2 TTS. Compared with ElevenLabs, OpenAI TTS, and Azure Speech, it is useful for reference-free character voice design, ad voiceovers, short-video narration, and multi-style voice exploration.

$0/1M tokens
XiaomiMiMoXiaomi MiMoMimo V2.5 TTS Voiceclone

Built for few-sample voice cloning, with stronger V2.5-series high-fidelity timbre reproduction, voice-character consistency, and cross-text generalization than MiMo V2 TTS. Compared with ElevenLabs, OpenAI TTS, and Azure Speech, it is well suited to Chinese character dubbing, short-video narration, podcast voice replication, and multilingual speech production that needs to preserve a target voice.

$0/1M tokens
XiaomiMiMoXiaomi MiMoMimo V2.5 TTS

Built for text-to-speech, voice design, and voice cloning, with stronger V2.5-series natural speech synthesis, fine-grained pace/emotion/tone control, and low-friction voice creation than MiMo V2 TTS. Compared with ElevenLabs, OpenAI TTS, and Azure Speech, it is well suited to Chinese narration, character dubbing, short-video audio, and multilingual voice production.

$0/1M tokens

FAQ

Frequently asked questions

What is Mimo V2.5 best suited for?

Mimo V2.5 is suited to the input, output, and capability types listed on this page. Test production workloads before choosing it for a critical system.

How is Mimo V2.5 priced?

Pricing is shown in the pricing section above. Actual cost depends on usage volume, account group, and the request payload.

How can I call Mimo V2.5?

Use the model ID shown above with one of the supported API protocols. Code examples are provided when usage data is available.

Get started

Build with Mimo V2.5

Use in Console
TokenHot

The frontier intelligence gateway. One API. 127 models. 0.2s latency. Pay only for what you use.

All systems normal · 99.997% uptime

Company

© 2026 TokenHot Inc. — Built for builders.