Qwen3.8-Max is the new Tongyi Qianwen flagship officially released by Alibaba on August 3, 2026: a MoE model with 2.4T total parameters, 95B active, a 1M-token context window, native multimodality, and official claims of comprehensive upgrades in coding and professional cowork (source). Its price also splits into two paths: Qwen AI Token Plan membership (personal from ¥39/mo, team from ¥150/seat/mo) and Tongyi Qianwen API pay-as-you-go (Qwen3.8-Max ¥12 input / ¥36 output per 1M tokens) — two completely different pricing models whose numbers cannot be compared directly (membership pricing, API pricing). All prices below are as of the official pages on 2026-08-18.
Qwen3.8-Max Release Overview: The 2.4T New Flagship Goes Live
Qwen3.8-Max officially launched on August 3, 2026 as the largest model in Tongyi Qianwen history, with weights to be open-sourced next week (including Qwen3.8-27B) (source). It first appeared as a Preview on July 19, 2026, then went fully live with API access on August 3, positioned as a "coding + cowork" dual-core flagship built for long-horizon autonomous tasks — in official demos it ran autonomously for over 10 days to deliver an open-source project end to end from an empty directory.
Key specs:
| Spec | Value |
|---|---|
| Total parameters | 2.4T (MoE sparse activation) |
| Active parameters | ~95B |
| Context window | 1M tokens |
| Input modalities | Text, image, video, document |
| Release timeline | Preview 2026-07-19 → Official 2026-08-03 |
| Open-source plan | Weights open next week, Qwen3.8-27B alongside |
For users, Qwen3.8-Max means you can either use the flagship inside Qwen AI Token Plan or call it pay-as-you-go on the Tongyi Qianwen API — the two paths differ in entry, top-up, quota, and rate-limit rules, compared item by item below.
Qwen3.8-Max Pricing Comparison: Membership and API Are Two Separate Systems
You cannot look up a single number for Qwen3.8-Max pricing — Qwen AI Token Plan membership and Tongyi Qianwen API are two fully independent billing systems, and membership Credits do not equal API balance. Token Plan is subscribed at platform.qianwenai.com, paying monthly/quarterly/yearly for a Credits pool; the API is prepaid on Alibaba Cloud Bailian, deducting balance by actual token consumption. High-level comparison:
| Dimension | Qwen AI Token Plan | Tongyi Qianwen API |
|---|---|---|
| Billing model | Subscription; monthly/quarterly/yearly for Credits | Pay-as-you-go; per-token deduction from balance |
| Official entry | platform.qianwenai.com/pricing/token-plan | help.aliyun.com/zh/model-studio/model-pricing |
| Top-up | Subscription charges; multiple cadences | Prepaid balance in the Bailian console |
| Getting Qwen3.8-Max | First-taste model, available on paid tiers | Just select the qwen3.8-max model |
| Entry price | ¥39/mo (personal Lite) | ¥12 input / ¥36 output per 1M tokens |
| Rate limits | Credits dual caps + agent concurrency | Per-model/account limits, per console |
| Best for | Interactive use inside compatible tools | Building apps, backend services, product integration |
Membership is a "quota pool" model: on the personal plans, 10 models (Qwen3.8-Max, Qwen3.7-Max, Qwen3.7-Plus, Qwen3.6-Flash, DeepSeek-V4-Pro, DeepSeek-V4-Flash, GLM-5.2, Wan2.7-Image, Wan2.7-Image-Pro, HappyHorse1.1) share one Credits pool, and Harness tools like web search, image-to-text, web scraping, and code interpreter all draw from the same allowance. The API is a "balance" model: you top up first, then every call settles by token count and cache-hit status. The clearest difference: membership resets quota monthly at a fixed price; the API charges exactly what you use, so cost grows linearly with usage.
Qwen Token Plan Membership Pricing: Three Personal Tiers and Qwen3.8-Max
All three Qwen AI Token Plan personal tiers (Lite / Standard / Pro) can call Qwen3.8-Max from ¥39/mo, distinguished by Credits quota and agent concurrency, with multi-cadence billing saving more; the team edition is sold per seat, with the Standard seat from ¥150/seat/mo. Pricing and quota per tier (source):
| Tier | Monthly (regular) | Quarterly | Annual | Credits/7 days | Agent concurrency |
|---|---|---|---|---|---|
| Lite | ¥39 (¥60) | ¥110/qtr | ¥420/yr | 2,500 + 700/5h | 1-2 |
| Standard | ¥139 (¥180) | ¥396/qtr | ¥1,510/yr | 10,000 + 3,000/5h | 3-4 |
| Pro | ¥499 (¥600) | ¥1,420/qtr | ¥5,600/yr | 40,000 + 12,000/5h | 6-8 |
| Team Standard seat | ¥150 (¥198) | — | ¥1,800/seat/yr | 25,000/seat/mo | — |
| Team Advanced seat | ¥550 (¥698) | — | ¥6,600/seat/yr | 100,000/seat/mo | — |
| Team Premier seat | ¥1,398 | — | ¥16,776/seat/yr | 250,000/seat/mo | — |
How to pick: for light personal trials of Qwen3.8-Max, Lite at ¥39/mo is enough; for daily coding, multi-file refactoring, and medium agent workloads, Standard ¥139/mo (4x Credits); for heavy professional use with lots of multimodal output and multi-agent concurrency, Pro ¥499/mo (16x Credits). For teams, choose Standard/Advanced/Premier seats by size and usage — the team edition commits that data is not used for model optimization. All tiers enjoy lower Credits consumption at night (00:00-08:00), so scheduling non-urgent tasks overnight stretches your quota.
Tongyi Qianwen API Pricing Comparison: Qwen3.8-Max vs Other Models
On the Tongyi Qianwen API, Qwen3.8-Max costs ¥12 input / ¥36 output per 1M tokens (0<Token≤1M tier), the same standard price as Qwen3.7-Max; Qwen3.7-Max's main ID is currently 50% off (¥6/¥18), making it the budget flagship alternative. Common model pricing (source):
| Model | Input (per 1M tokens) | Output | Notes |
|---|---|---|---|
| Qwen3.8-Max | ¥12 | ¥36 | New flagship, 1M context, coding/cowork dual-core |
| Qwen3.7-Max | ¥12 (limited-time ¥6) | ¥36 (limited-time ¥18) | Previous flagship, main ID 50% off |
| Qwen3.7-Plus | ¥2 (limited-time ¥1.6) | ¥8 (limited-time ¥6.4) | Production workhorse, reasoning/vision/text |
| Qwen3.6-Flash | ~¥1.2 | ~¥7.2 | Speed and cost focused |
| Qwen-Turbo | ¥0.3 | ¥0.6 | Cost-effective entry (non-thinking mode) |
On the API side, Qwen3.8-Max is a "critical path model": give it complex reasoning, long-context analysis, multi-step agents, and coding workflows, while routing lightweight high-frequency traffic to Turbo / Flash / Plus to save money. The most valuable API-side features are tiered billing, Batch 50% pricing, and context caching: some models are priced by the input-token range of a single request (e.g., 0<Token≤256K vs 256K<Token≤1M); Batch calls bill input and output at 50% of real-time prices; context caching discounts only input tokens, and Batch cannot stack with caching. New users usually get free quota with an expiry date that may not be shared across models — confirm your balance and free-package details in the console before going live.
Quota, Usage and Rate Limits: Two Different Accounting Models
Membership is about "Credits quota + dual caps," while the API is about "balance + per-model/account rate limits" — two completely different consumption models. On the membership side, personal plans combine a 7-day total Credits cap with a 5-hour continuous usage cap; when either cap is hit, service pauses and auto-recovers on the next cycle reset. Team seats track Credits independently. On the API side, quota is your account balance plus per-model free packages/resource packs, and rate limits usually vary by model and account — check the exact RPM/TPM numbers in the Bailian console.
Do not look only at unit price before launch: for high-concurrency API workloads, first confirm your account rate-limit thresholds can carry the target traffic; for membership, evaluate whether the Credits are enough — Pro's 40,000 Credits/7 days sustains all-day multi-agent tasks, while Lite's 2,500 Credits is closer to a light experience. Also note membership is limited to interactive use inside compatible tools — it cannot be used for automation scripts or application backends; backend integration requires the API.
API Key and Integration Config: Membership Toolchain vs Open API
Members do not need an open-platform API key: Qwen AI Token Plan uses a dedicated sk-sp- API Key with a dedicated Base URL for compatible tools such as Qwen Code, Claude Code, Cursor, Cline, OpenCode, Codex, Kilo CLI, and OpenClaw, used interactively within your Credits quota; for building your own app or product integration, you apply for a Tongyi Qianwen API key in the Alibaba Cloud Bailian console and integrate into your backend in the OpenAI-compatible format. The two keys are fully independent and cannot be interchanged — this is the most common misconception.
Setup differs by scenario: coding-tool users install Qwen Code / Claude Code / Cline and select the Qwen Token Plan channel, then fill in the sk-sp- Key and Base URL; backend developers create an API key in the Bailian console, configure the base URL and model name (e.g., qwen3.8-max) in the OpenAI-compatible format, and enable Batch or context caching as needed. Qwen3.8-Max supports native tool calls (code_interpreter, web_search, etc.), making it suitable for complex agent pipelines.
Deals and Is It Worth It: How to Use Qwen3.8-Max Smartly
Bottom line: if you work inside Qwen AI Token Plan, subscribe to membership — light users start at Lite ¥39/mo, heavy coding users go Standard/Pro, and multi-cadence billing plus night discounts spread the cost; if you build your own app, the Tongyi Qianwen API is the only option, where Qwen3.8-Max's unit price matches 3.7-Max at standard rates, but Batch 50% pricing, cache discounts, and the Qwen3.7-Max limited-time 50% off help control costs. Membership deals are mainly multi-cadence discounts (annual is cheapest) and lower night Credits consumption; API deals are Batch 50% pricing, context-cache discounts, and free quota for new users.
| Your situation | Recommendation |
|---|---|
| Light Qwen3.8-Max trial, first taste | Qwen Token Plan personal Lite ¥39/mo |
| Daily coding + agents + office mix | Standard ¥139/mo (¥1,510/yr) |
| Heavy multi-agent, lots of multimodal output | Pro ¥499/mo (¥5,600/yr) |
| Team collaboration, data not used for training | Team Standard seat from ¥150/seat/mo |
| Building apps, product integration, backend services | Tongyi Qianwen API, Qwen3.8-Max pay-as-you-go ¥12/¥36 |
For broader comparisons against other Chinese models, check the DeepSeek API and Zhipu GLM comparison pages. Whether Qwen3.8-Max is worth it depends on which path you take: membership trades a monthly budget for a fixed capability pool, while the API trades per-use fees for flexibility — pick the right entry point first, then compare the prices.