Mimo Token API Pricing, Top-up, Deals & Updates | Article

Mimo Token API pricing. mimo-v2.5-pro, mimo-v2.5 and 4 more models. OpenAI API, Anthropic API, MiMo API integrations. Includes cache tiers and rate limits. API key setup, usage quotas, and official entry—is it worth it?

41
Platforms
395+
Models
140+
IDE Tools
6
Mimo Token API models
¥1 / ¥2
Mimo Token API entry price
View
Vendor details

MiMo API Latest Updates

Xiaomi MiMo-V2.5 API Price Cut Up to 99%: Pro ¥3/¥6 vs ¥7/¥21 Pricing Comparison

Published Updated
PricingMiMoAPIPricing ComparisonPrice CutV2.5Xiaomi
Summary

Xiaomi MiMo-V2.5 API pricing was permanently cut on May 27, 2026, by up to 99%, removing context-length tiers. MiMo-V2.5-Pro now costs ¥0.025 cache-hit / ¥3 input / ¥6 output per 1M tokens, and the standard model ¥0.02 / ¥1 / ¥2, versus the old V2 series (Pro ¥7/¥21). This pricing comparison breaks down the real reduction on cache-hit input, cache-miss input and output, and compares MiMo against DeepSeek and Claude to help you decide if it is worth switching.

How much the price actually dropped

Xiaomi MiMo-V2.5 API pricing was permanently cut starting 00:00 Beijing time on May 27, 2026, by up to 99%, and the context-length tiering was removed. The official announcement frames it as "a permanent overhaul of the entire model pricing system," applied globally at once (source).

The direct effect for you: the same call volume now produces a much smaller bill, especially on cache hits and long-context workloads.

Before-and-after pricing comparison

The official pricing page lists the old V2 series and the new V2.5 series side by side, which makes a clean pricing comparison (source):

Model Billing item Old (V2 series) New (V2.5 series) Cut
Pro Cache-hit input ¥1.40 (≤256K) / ¥2.80 (256K-1M) ¥0.025 98%-99%
Pro Cache-miss input ¥7.00 / ¥14.00 ¥3.00 57%-79%
Pro Output ¥21.00 / ¥42.00 ¥6.00 71%-86%
Standard Cache-hit input ¥0.56 ¥0.02 96%
Standard Cache-miss input ¥2.80 ¥1.00 64%
Standard Output ¥14.00 ¥2.00 86%

All prices are CNY per 1M tokens. The old column is the V2 series (Pro maps to mimo-v2-pro, standard to mimo-v2-omni), which is the baseline behind the "up to 99%" claim.

The cut is not even: cache hits are the big one

The "up to 99%" refers to cache-hit input, not a 90% discount across every tier. Break it into three lines:

  • Cache-hit input: Pro ¥2.80 (long context) → ¥0.025, down 99%; nearly free input.
  • Cache-miss input: Pro ¥7 → ¥3, down 57%; long context ¥14 → ¥3, down 79%.
  • Output: Pro ¥21 → ¥6, down 71%; long context ¥42 → ¥6, down 86%.

So the real savings depend on your workload mix. The more repeated context you have (multi-turn chats, agents re-reading tool definitions, fixed RAG documents), the higher your cache-hit rate and the closer you get to that 99% line. Purely one-shot generation mostly captures the 71%-86% output reduction.

Pricing comparison against rivals: worth switching?

Put MiMo-V2.5 side by side with mainstream domestic and international models, and its value is strong. Using flagship cache-miss input/output as the benchmark:

Model Cache-miss input Output Cache-hit input
MiMo-V2.5-Pro ¥3.00 ¥6.00 ¥0.025
MiMo-V2.5 (standard) ¥1.00 ¥2.00 ¥0.02
DeepSeek V4 Pro ¥3.13 ¥6.26 ¥0.025-0.05
Claude Opus 4.5 ¥108 ¥540 ~¥15

This pricing comparison shows MiMo-V2.5-Pro is basically in the same tier as DeepSeek V4 Pro (with a slightly cheaper input) and matches it on cache-hit pricing, while undercutting Claude Opus 4.5 by roughly 97% on input. If long-context inference cost is your bottleneck, MiMo's 1M context window plus cheap cache hits are two concrete selling points. DeepSeek and Claude figures are list prices with slightly different cache-hit definitions; check each provider's pricing page for exact numbers.

For existing users: quota reset + 5-8x volume

If you already use a MiMo Token Plan, this was not just an API price cut — the plan also got more generous. Xiaomi optimized the Token Plan billing system so the same price now yields 5-8x the usable token volume, and all active subscribers had their Credits reset at 00:00 on May 27 under the new rules (source).

Keep the two entries separate: the open API uses sk- keys and deducts account balance, while Token Plan uses tp- keys against a Credits pool — fully independent billing systems. Use MiMo API pay-as-you-go for backend services and custom apps; use MiMo Token Plan for a fixed monthly coding subscription.

Things to watch

  1. The old V2 series was retired on June 30, 2026. Legacy names like mimo-v2-pro and mimo-v2-omni auto-routed to V2.5 equivalents from June 1 and went offline on June 30; new integrations should use V2.5 directly.

  2. The cut removed input-length tiering, not every add-on. Web search (¥16/1K calls domestically) and ASR (¥0.5/hour) are still billed separately from token pricing; the TTS family is temporarily free.

  3. Cache hits are the main cost lever. Fixed system prompts, repeated project context, and shared multi-turn history reliably capture the ¥0.025/¥0.02 cache-hit rate, pushing input cost close to free.

  4. Overseas pricing is in USD. Pro maps to $0.0036/$0.435/$0.87 and the standard model to $0.0028/$0.14/$0.28 per 1M tokens, adjusted globally at the same time.

Sources and verification

Last verified

Mimo Token API FAQ

How much did the Xiaomi MiMo-V2.5 API price cut actually drop, and where does the 99% come from?

The 99% figure is the cache-hit input tier: Pro drops from ¥1.40 (≤256K) to ¥0.025 per 1M tokens, and from ¥2.80 (256K-1M) to ¥0.025. Cache-miss input drops from ¥7 to ¥3 (57% off) and output from ¥21 to ¥6 (71% off); long-context output falls from ¥42 to ¥6 (86% off). So 99% applies to cache hits, not every billing tier.

Does MiMo-V2.5 still charge more for longer context after the price cut?

No. Since May 27 the MiMo-V2.5 series no longer splits pricing by ≤256K / 256K-1M input length; Pro and the standard model use flat rates. The old V2 rule that doubled input and output above 256K is gone, so long-context calls no longer jump into a higher price band.

How does MiMo-V2.5 compare with DeepSeek and Claude, and is it worth switching?

It is worth a close look. MiMo-V2.5-Pro cache-miss input ¥3/output ¥6 sits in the same tier as DeepSeek V4 Pro (¥3.13/¥6.26) with a slightly cheaper input, and its ¥0.025 cache-hit price is very competitive; against Claude Opus 4.5 (¥108/¥540) it is roughly two orders of magnitude cheaper. For agent and long-context reasoning, MiMo's 1M context plus cheap cache hits is a clear advantage.

Did the MiMo Token Plan subscription change along with the API price cut?

Yes. Xiaomi also optimized the Token Plan billing system, boosting usable token volume by 5-8x at the same price, and reset all active subscribers' Credits quotas on May 27. If you already use a Token Plan, your quota is recalculated under the new rules automatically.

How does MiMo API context caching trigger after the cut, and how much can it save?

When a request prefix (system prompt, chat history, RAG documents) hits the Prompt Cache, it is billed at the cache-hit rate automatically, with no developer config and free cache writes. A hit drops Pro input from ¥3 to ¥0.025 (about 99% saved) and standard input from ¥1 to ¥0.02 (about 98% saved), which benefits multi-turn chats and repeated project contexts the most.
AI Token Plan

Global AI Token Comparison

AI Token Plan is not just a collection of vendor links. It places 41+ domestic and international AI platforms into one comparison framework, covering 395+ model entries plus common plan types such as Token Plans, Coding Plans, IDE Tools and LLM APIs.

When you need to compare AI subscription pricing, coding plan quota, API usage cost or official deal entry points, AI Token Plan brings official prices, plan tiers, usage rules, model capabilities and tool integrations into one place, reducing the need to check multiple vendor sites manually.

Data is continuously organized as vendor pricing pages, plan pages and product documentation change, making the homepage a pricing comparison entry point while detail pages explain whether each vendor plan fits individual developers, team purchasing or long-term API usage.

Prices come from each platform's official site and may change at any time; the official price prevails.

© 2026 AI Token Plan · All rights reserved · First published June 18, 2026 · 64 days running · Sitemap