Qianwen Token Plan Pricing, Top-up, Deals & Updates | Article

Qianwen Token Plan subscription pricing. 6 tiers (Token Plan (Individual/Team), from ¥39/mo/mo). with 2,500-250,000 Credits. Qwen3.8-Max and 9 more models. Qwen Code, Claude Code, Cursor and 5 more tools integrations. Plan quotas, perks, and official entry—is it worth it?

41
Platforms
395+
Models
140+
IDE Tools
10
Qianwen Token Plan models
¥39/mo
Qianwen Token Plan entry price
View
Vendor details

Qianwen AI Latest Updates

Qwen3.8-Max Release Pricing Comparison: Qwen Token Plan Membership vs Tongyi Qianwen API Pay-As-You-Go — Top-up, Deals, Quota, Rate Limits and Is It Worth It

Published Updated
ModelQwen3.8-MaxQwen AIToken Planpricing comparisonmodel releaseTongyi Qianwen API
Summary

Qwen3.8-Max officially launched on August 3, 2026: a 2.4T-parameter flagship with 95B active parameters, priced at ¥12 input / ¥36 output per 1M tokens on the API, the same as Qwen3.7-Max. On Qwen Token Plan, the personal Lite plan starts at ¥39/mo and the team Standard seat at ¥150/seat/mo. This comparison covers Qwen3.8-Max pricing across membership vs API pay-as-you-go — top-up, deals, official access, quota, usage, rate limits, API key setup and integration config — to help you decide whether it is worth it.

Qwen3.8-Max is the new Tongyi Qianwen flagship officially released by Alibaba on August 3, 2026: a MoE model with 2.4T total parameters, 95B active, a 1M-token context window, native multimodality, and official claims of comprehensive upgrades in coding and professional cowork (source). Its price also splits into two paths: Qwen AI Token Plan membership (personal from ¥39/mo, team from ¥150/seat/mo) and Tongyi Qianwen API pay-as-you-go (Qwen3.8-Max ¥12 input / ¥36 output per 1M tokens) — two completely different pricing models whose numbers cannot be compared directly (membership pricing, API pricing). All prices below are as of the official pages on 2026-08-18.

Qwen3.8-Max Release Overview: The 2.4T New Flagship Goes Live

Qwen3.8-Max officially launched on August 3, 2026 as the largest model in Tongyi Qianwen history, with weights to be open-sourced next week (including Qwen3.8-27B) (source). It first appeared as a Preview on July 19, 2026, then went fully live with API access on August 3, positioned as a "coding + cowork" dual-core flagship built for long-horizon autonomous tasks — in official demos it ran autonomously for over 10 days to deliver an open-source project end to end from an empty directory.

Key specs:

Spec Value
Total parameters 2.4T (MoE sparse activation)
Active parameters ~95B
Context window 1M tokens
Input modalities Text, image, video, document
Release timeline Preview 2026-07-19 → Official 2026-08-03
Open-source plan Weights open next week, Qwen3.8-27B alongside

For users, Qwen3.8-Max means you can either use the flagship inside Qwen AI Token Plan or call it pay-as-you-go on the Tongyi Qianwen API — the two paths differ in entry, top-up, quota, and rate-limit rules, compared item by item below.

Qwen3.8-Max Pricing Comparison: Membership and API Are Two Separate Systems

You cannot look up a single number for Qwen3.8-Max pricing — Qwen AI Token Plan membership and Tongyi Qianwen API are two fully independent billing systems, and membership Credits do not equal API balance. Token Plan is subscribed at platform.qianwenai.com, paying monthly/quarterly/yearly for a Credits pool; the API is prepaid on Alibaba Cloud Bailian, deducting balance by actual token consumption. High-level comparison:

Dimension Qwen AI Token Plan Tongyi Qianwen API
Billing model Subscription; monthly/quarterly/yearly for Credits Pay-as-you-go; per-token deduction from balance
Official entry platform.qianwenai.com/pricing/token-plan help.aliyun.com/zh/model-studio/model-pricing
Top-up Subscription charges; multiple cadences Prepaid balance in the Bailian console
Getting Qwen3.8-Max First-taste model, available on paid tiers Just select the qwen3.8-max model
Entry price ¥39/mo (personal Lite) ¥12 input / ¥36 output per 1M tokens
Rate limits Credits dual caps + agent concurrency Per-model/account limits, per console
Best for Interactive use inside compatible tools Building apps, backend services, product integration

Membership is a "quota pool" model: on the personal plans, 10 models (Qwen3.8-Max, Qwen3.7-Max, Qwen3.7-Plus, Qwen3.6-Flash, DeepSeek-V4-Pro, DeepSeek-V4-Flash, GLM-5.2, Wan2.7-Image, Wan2.7-Image-Pro, HappyHorse1.1) share one Credits pool, and Harness tools like web search, image-to-text, web scraping, and code interpreter all draw from the same allowance. The API is a "balance" model: you top up first, then every call settles by token count and cache-hit status. The clearest difference: membership resets quota monthly at a fixed price; the API charges exactly what you use, so cost grows linearly with usage.

Qwen Token Plan Membership Pricing: Three Personal Tiers and Qwen3.8-Max

All three Qwen AI Token Plan personal tiers (Lite / Standard / Pro) can call Qwen3.8-Max from ¥39/mo, distinguished by Credits quota and agent concurrency, with multi-cadence billing saving more; the team edition is sold per seat, with the Standard seat from ¥150/seat/mo. Pricing and quota per tier (source):

Tier Monthly (regular) Quarterly Annual Credits/7 days Agent concurrency
Lite ¥39 (¥60) ¥110/qtr ¥420/yr 2,500 + 700/5h 1-2
Standard ¥139 (¥180) ¥396/qtr ¥1,510/yr 10,000 + 3,000/5h 3-4
Pro ¥499 (¥600) ¥1,420/qtr ¥5,600/yr 40,000 + 12,000/5h 6-8
Team Standard seat ¥150 (¥198) ¥1,800/seat/yr 25,000/seat/mo
Team Advanced seat ¥550 (¥698) ¥6,600/seat/yr 100,000/seat/mo
Team Premier seat ¥1,398 ¥16,776/seat/yr 250,000/seat/mo

How to pick: for light personal trials of Qwen3.8-Max, Lite at ¥39/mo is enough; for daily coding, multi-file refactoring, and medium agent workloads, Standard ¥139/mo (4x Credits); for heavy professional use with lots of multimodal output and multi-agent concurrency, Pro ¥499/mo (16x Credits). For teams, choose Standard/Advanced/Premier seats by size and usage — the team edition commits that data is not used for model optimization. All tiers enjoy lower Credits consumption at night (00:00-08:00), so scheduling non-urgent tasks overnight stretches your quota.

Tongyi Qianwen API Pricing Comparison: Qwen3.8-Max vs Other Models

On the Tongyi Qianwen API, Qwen3.8-Max costs ¥12 input / ¥36 output per 1M tokens (0<Token≤1M tier), the same standard price as Qwen3.7-Max; Qwen3.7-Max's main ID is currently 50% off (¥6/¥18), making it the budget flagship alternative. Common model pricing (source):

Model Input (per 1M tokens) Output Notes
Qwen3.8-Max ¥12 ¥36 New flagship, 1M context, coding/cowork dual-core
Qwen3.7-Max ¥12 (limited-time ¥6) ¥36 (limited-time ¥18) Previous flagship, main ID 50% off
Qwen3.7-Plus ¥2 (limited-time ¥1.6) ¥8 (limited-time ¥6.4) Production workhorse, reasoning/vision/text
Qwen3.6-Flash ~¥1.2 ~¥7.2 Speed and cost focused
Qwen-Turbo ¥0.3 ¥0.6 Cost-effective entry (non-thinking mode)

On the API side, Qwen3.8-Max is a "critical path model": give it complex reasoning, long-context analysis, multi-step agents, and coding workflows, while routing lightweight high-frequency traffic to Turbo / Flash / Plus to save money. The most valuable API-side features are tiered billing, Batch 50% pricing, and context caching: some models are priced by the input-token range of a single request (e.g., 0<Token≤256K vs 256K<Token≤1M); Batch calls bill input and output at 50% of real-time prices; context caching discounts only input tokens, and Batch cannot stack with caching. New users usually get free quota with an expiry date that may not be shared across models — confirm your balance and free-package details in the console before going live.

Quota, Usage and Rate Limits: Two Different Accounting Models

Membership is about "Credits quota + dual caps," while the API is about "balance + per-model/account rate limits" — two completely different consumption models. On the membership side, personal plans combine a 7-day total Credits cap with a 5-hour continuous usage cap; when either cap is hit, service pauses and auto-recovers on the next cycle reset. Team seats track Credits independently. On the API side, quota is your account balance plus per-model free packages/resource packs, and rate limits usually vary by model and account — check the exact RPM/TPM numbers in the Bailian console.

Do not look only at unit price before launch: for high-concurrency API workloads, first confirm your account rate-limit thresholds can carry the target traffic; for membership, evaluate whether the Credits are enough — Pro's 40,000 Credits/7 days sustains all-day multi-agent tasks, while Lite's 2,500 Credits is closer to a light experience. Also note membership is limited to interactive use inside compatible tools — it cannot be used for automation scripts or application backends; backend integration requires the API.

API Key and Integration Config: Membership Toolchain vs Open API

Members do not need an open-platform API key: Qwen AI Token Plan uses a dedicated sk-sp- API Key with a dedicated Base URL for compatible tools such as Qwen Code, Claude Code, Cursor, Cline, OpenCode, Codex, Kilo CLI, and OpenClaw, used interactively within your Credits quota; for building your own app or product integration, you apply for a Tongyi Qianwen API key in the Alibaba Cloud Bailian console and integrate into your backend in the OpenAI-compatible format. The two keys are fully independent and cannot be interchanged — this is the most common misconception.

Setup differs by scenario: coding-tool users install Qwen Code / Claude Code / Cline and select the Qwen Token Plan channel, then fill in the sk-sp- Key and Base URL; backend developers create an API key in the Bailian console, configure the base URL and model name (e.g., qwen3.8-max) in the OpenAI-compatible format, and enable Batch or context caching as needed. Qwen3.8-Max supports native tool calls (code_interpreter, web_search, etc.), making it suitable for complex agent pipelines.

Deals and Is It Worth It: How to Use Qwen3.8-Max Smartly

Bottom line: if you work inside Qwen AI Token Plan, subscribe to membership — light users start at Lite ¥39/mo, heavy coding users go Standard/Pro, and multi-cadence billing plus night discounts spread the cost; if you build your own app, the Tongyi Qianwen API is the only option, where Qwen3.8-Max's unit price matches 3.7-Max at standard rates, but Batch 50% pricing, cache discounts, and the Qwen3.7-Max limited-time 50% off help control costs. Membership deals are mainly multi-cadence discounts (annual is cheapest) and lower night Credits consumption; API deals are Batch 50% pricing, context-cache discounts, and free quota for new users.

Your situation Recommendation
Light Qwen3.8-Max trial, first taste Qwen Token Plan personal Lite ¥39/mo
Daily coding + agents + office mix Standard ¥139/mo (¥1,510/yr)
Heavy multi-agent, lots of multimodal output Pro ¥499/mo (¥5,600/yr)
Team collaboration, data not used for training Team Standard seat from ¥150/seat/mo
Building apps, product integration, backend services Tongyi Qianwen API, Qwen3.8-Max pay-as-you-go ¥12/¥36

For broader comparisons against other Chinese models, check the DeepSeek API and Zhipu GLM comparison pages. Whether Qwen3.8-Max is worth it depends on which path you take: membership trades a monthly budget for a fixed capability pool, while the API trades per-use fees for flexibility — pick the right entry point first, then compare the prices.

Sources and verification

Last verified

Qianwen Token Plan FAQ

Qwen3.8-Max pricing comparison: Qwen Token Plan membership vs API pay-as-you-go — which is better?

There is no absolute answer: to use Qwen3.8-Max inside Qwen AI Token Plan with a shared Credits pool across models, the personal Lite plan at ¥39/mo is enough to try it; for building your own app, product integration, or server-side calls you must use the Tongyi Qianwen API pay-as-you-go (Qwen3.8-Max ¥12 input / ¥36 output per 1M tokens). Membership resets quota monthly at a fixed price; the API charges exactly what you use — pick the right entry point before comparing prices.

Is Qwen3.8-Max API priced the same as Qwen3.7-Max, and which is worth using?

Qwen3.8-Max and Qwen3.7-Max share the same standard API price (¥12 input / ¥36 output per 1M tokens on the 0<Token≤1M tier), but Qwen3.7-Max main ID is currently 50% off (¥6/¥18). Qwen3.8-Max is the new 2.4T flagship with 1M-token context and stronger coding and cowork capabilities — pick 3.8-Max for top capability, or take advantage of the 3.7-Max discount to control costs.

Which Qwen Token Plan tier can use Qwen3.8-Max, and is upgrading the personal plans worth it?

Qwen3.8-Max is a first-taste model on Qwen AI Token Plan (from 2026-07), and all three personal tiers — Lite / Standard / Pro — can use it, differing only in Credits quota: Lite 2,500 Credits/7 days, Standard 10,000, Pro 40,000. Lite is enough for light daily use; Standard ¥139/mo suits frequent agents or heavier coding; Pro ¥499/mo fits heavy multimodal output and multi-agent concurrency.

What is the difference in API key and integration config between Qwen Token Plan and Tongyi Qianwen API for Qwen3.8-Max?

Qwen Token Plan uses a dedicated sk-sp- API Key with a dedicated Base URL, limited to interactive use inside compatible tools (Qwen Code, Claude Code, Cursor, Cline, etc.) — not for automation or backends. The Tongyi Qianwen API requires a separate key from the Alibaba Cloud Bailian console, integrated into backends via the OpenAI-compatible format. The two keys are fully independent and cannot be interchanged.

Is Qwen3.8-Max worth it, and what should I watch out for with quota, rate limits, and night discounts?

Qwen3.8-Max is worth it for long-horizon coding, complex agents, and long-context analysis on a controlled budget — on Qwen Token Plan it shares one Credits pool with 10 models including DeepSeek-V4-Pro and GLM-5.2, with lower Credits consumption at night (00:00-08:00). On the API side, note tiered billing and that Batch 50% pricing cannot stack with context caching. Confirm your account rate-limit tier and quota before production.
AI Token Plan

Global AI Token Comparison

AI Token Plan is not just a collection of vendor links. It places 41+ domestic and international AI platforms into one comparison framework, covering 395+ model entries plus common plan types such as Token Plans, Coding Plans, IDE Tools and LLM APIs.

When you need to compare AI subscription pricing, coding plan quota, API usage cost or official deal entry points, AI Token Plan brings official prices, plan tiers, usage rules, model capabilities and tool integrations into one place, reducing the need to check multiple vendor sites manually.

Data is continuously organized as vendor pricing pages, plan pages and product documentation change, making the homepage a pricing comparison entry point while detail pages explain whether each vendor plan fits individual developers, team purchasing or long-term API usage.

Prices come from each platform's official site and may change at any time; the official price prevails.

© 2026 AI Token Plan · All rights reserved · First published June 18, 2026 · 64 days running · Sitemap