Deepseek Token API Pricing, Top-up, Deals & Updates | Article

Deepseek Token API pricing. DeepSeek-V4-Flash, DeepSeek-V4-Pro and 1 more models. OpenAI-compatible API, Anthropic-compatible API, Claude Code and 3 more tools integrations. Includes cache tiers and rate limits. API key setup, usage quotas, and official entry—is it worth it?

41
Platforms
395+
Models
140+
IDE Tools
3
Deepseek Token API models
¥1.5–3 / ¥4.5–9
Deepseek Token API entry price
View
Vendor details

DeepSeek API Latest Updates

DeepSeek V4-Pro Official Release: $0.435 Input / $0.87 Output Per MTok, What's Different From the Preview

Published Updated
ModelDeepSeekV4-ProOfficial ReleaseModel ReleaseAPIPricing
Summary

DeepSeek V4-Pro is officially released as DeepSeek-V4-Pro-0813, graduating from preview with enhanced Agent capability and Responses API + Codex integration. Pricing is unchanged: $0.435/MTok input (cache miss), $0.87/MTok output, $0.003625/MTok cache-hit, concurrency 500, model name still deepseek-v4-pro. Here's what changed and how it compares to V4-Flash.

DeepSeek V4-Pro is officially released, now versioned DeepSeek-V4-Pro-0813 — with pricing, specs, and the API call name all unchanged. It's a "preview to official" graduation: both models in the V4 lineup are now fully official (source).

What Actually Changed

V4-Pro graduated from preview, and the official pricing page now lists DeepSeek-V4-Pro-0813. The July 31 changelog entry already said the V4-Pro official release "will follow soon" (source). The version bump on the pricing page means it's now live. That's the whole story: the version number changed, nothing else did.

Item Before Now (official)
Model version V4-Pro preview DeepSeek-V4-Pro-0813
API call name deepseek-v4-pro deepseek-v4-pro
Input (cache miss) $0.435 / MTok $0.435 / MTok
Output $0.87 / MTok $0.87 / MTok
Input (cache hit) $0.003625 / MTok $0.003625 / MTok
Concurrency 500 500

Enhanced Agent Capability + Codex Integration

DeepSeek confirmed the V4-Pro official release enhances Agent capability and adds Responses API and Codex integration. The official announcement from DeepSeek's user community reads: "V4-Pro official release is now live on the API, model name unchanged, with enhanced Agent capability and Responses API + Codex integration — welcome to test and provide feedback." Responses API support is already visible on the pricing page (source), and Codex integration is now in the official docs — both deepseek-v4-pro and deepseek-v4-flash work directly in Codex CLI, the ChatGPT desktop app, and the VS Code Codex extension (source).

As for how much the Agent capability actually improved, DeepSeek hasn't published benchmark numbers yet, and the changelog hasn't caught up to this Pro entry. So the confirmed story is "capability up + Codex works" — the specific scores still need official data.

Pricing and Specs: Nothing Moved

This update involves no price change — pricing and specs are identical to before. Here's the current config for both models on the official pricing page (source):

Item V4-Flash V4-Pro
Input (cache hit) $0.0028 / MTok $0.003625 / MTok
Input (cache miss) $0.14 / MTok $0.435 / MTok
Output $0.28 / MTok $0.87 / MTok
Concurrency 2500 500
Context 1M 1M
Max output 384K 384K

Features are identical across both models too: thinking/non-thinking modes, JSON Output, Tool Calls, Responses API, and Anthropic API all supported; FIM completion is non-thinking mode only.

How to Choose Between V4-Flash and V4-Pro

Daily work goes to Flash, complex reasoning goes to Pro — that advice hasn't changed. Pro costs 3x more and has 5x lower concurrency, but what you're buying is stronger deep-reasoning capability. If your workload is complex refactors, long multi-step agent chains, or knowledge-intensive workflows where every step depends on the last, the 3x premium is worth it. Otherwise — daily coding, batch jobs, light agents — Flash is plenty.

Things to Watch

  1. No official benchmarks yet. The changelog still stops at the July 31 Flash entry. Pro's official technical report and agent benchmarks aren't out, so don't compare against the old preview scores.
  2. The price-hike warning is still live. The official pricing page explicitly says DeepSeek "plan[s] to raise the overall pricing ... with a significant increase expected." Get your caching strategy and usage estimates sorted before it hits.
  3. No code changes needed. Existing deepseek-v4-pro calls route to the new version automatically — no code changes, no key rotation.
  4. Prefer a fixed monthly fee over pay-as-you-go? Check out third-party plans from Volcengine and CtCloud, which aren't subject to the official concurrency cap or price hike.

Sources and verification

Last verified

Deepseek Token API FAQ

What's the difference between the official DeepSeek V4-Pro and the preview, and what does the 0813 version mean?

From the official pricing page, the main change is the version number: the model graduated from preview to DeepSeek-V4-Pro-0813. Pricing, specs, and the API call name are all unchanged — $0.435/MTok input (cache miss), $0.87/MTok output, $0.003625/MTok cache-hit, concurrency 500. DeepSeek also confirmed the new version enhances Agent capability and adds Responses API and Codex integration. It's a seamless server-side upgrade; your existing deepseek-v4-pro calls route to the new version automatically.

Did DeepSeek V4-Pro pricing change with the official release, and what are the input/output rates and concurrency?

No. Pricing is unchanged: $0.435/MTok input (cache miss), $0.87/MTok output, $0.003625/MTok cache-hit input, concurrency 500. Versus V4-Flash at $0.14/$0.28 and 2500 concurrency, Pro is 3x the price and 5x lower concurrency. Note the official page still carries a warning about a significant price hike coming soon — get your usage sorted before it lands.

I'm already using deepseek-v4-pro — do I need to change code, rotate API keys, or switch model names after the V4-Pro official release?

No. The model name stays deepseek-v4-pro, and your API key and BASE URL (https://api.deepseek.com) don't change. This is a server-side upgrade: existing requests route to DeepSeek-V4-Pro-0813 automatically with no client changes. The only thing to watch is the upcoming price hike — that will hit your bill, but it's separate from this graduation.

How do I choose between DeepSeek V4-Pro and V4-Flash, and what tasks justify paying 3x more for Pro?

Daily coding, batch jobs, and light agents belong on V4-Flash. V4-Pro is for complex refactors, long multi-step agent chains, and knowledge-intensive production workloads where every step depends on the previous one's reasoning quality. Also mind concurrency: Pro only has 500, so high-volume workloads should run Flash or use a third-party fixed plan.

Has DeepSeek published the official V4-Pro agent benchmarks, and what should I watch for next?

Not yet. The changelog still stops at the July 31 Flash entry, and Pro's official technical report and agent benchmarks haven't been released — don't compare against the old preview scores. Two things to watch: DeepSeek publishing the Pro official benchmarks, and the formal price-hike notice, since the pricing page already warns of a 'significant increase.'
AI Token Plan

Global AI Token Comparison

AI Token Plan is not just a collection of vendor links. It places 41+ domestic and international AI platforms into one comparison framework, covering 395+ model entries plus common plan types such as Token Plans, Coding Plans, IDE Tools and LLM APIs.

When you need to compare AI subscription pricing, coding plan quota, API usage cost or official deal entry points, AI Token Plan brings official prices, plan tiers, usage rules, model capabilities and tool integrations into one place, reducing the need to check multiple vendor sites manually.

Data is continuously organized as vendor pricing pages, plan pages and product documentation change, making the homepage a pricing comparison entry point while detail pages explain whether each vendor plan fits individual developers, team purchasing or long-term API usage.

Prices come from each platform's official site and may change at any time; the official price prevails.

© 2026 AI Token Plan · All rights reserved · First published June 18, 2026 · 64 days running · Sitemap