Global AI Token Plan comparison home

Baidu Token API Pricing, Top-up, Deals & Official Entry

This ERNIE API page compares Baidu Qianfan official API pricing, top-up, deals and API key application rules: ERNIE 5.1, ERNIE 5.0, ERNIE X1, Turbo, Lite Pro and Qianfan vision/OCR start from ¥0.2/¥0.4 per 1M tokens, while long context, image, OCR, web search and prepaid packs use different billing units. It brings OpenAI and Anthropic compatible APIs, ¥0.004 web search, prompt cache-hit pricing, Batch inference discounts, prepaid packs, rate limits and the Coding Plan dedicated-key boundary together, preventing regular API balance from being confused with coding subscription quota.

Token APIERNIE 5.1
SubscriptionToken API

Below is the complete pricing comparison for Baidu Token API, covering 4 plan tiers, core model capabilities, quotas, and official entry. All data is sourced from the official website to help you decide whether it’s worth it.

Last updated

Baidu Token API Price Comparison: Core Models

ERNIE-5.1ERNIE-5.0ERNIE-5.0-Thinking-PreviewERNIE-5.0-Thinking-LatestERNIE-5.0-Thinking-ExpERNIE-4.5-Turbo-128KERNIE-4.5-Turbo-128K-PreviewERNIE-4.5-Turbo-32KERNIE-4.5-Turbo-20260402ERNIE-4.5-Turbo-VLERNIE-4.5-Turbo-VL-32KERNIE-4.5-8KERNIE-4.5-0.3BERNIE-Speed-Pro-128KERNIE-Lite-Pro-128Kernie-char-8kernie-char-fiction-8kERNIE-X1.1-PreviewERNIE-X1.1Qianfan-Check-VLQianfan-VL-70BQianfan-VL-8BQianfan-VL-1.5-FlashQianfan-FuncCallerQianfan-ToyTalkQianfan-OCRMuseSteamer-Air-ImageMuseSteamer-Air-I2V
ERNIE-5.1

ERNIE 5.1 flagship, 128K context; ¥4 input, ¥18 output per 1M tokens (≤32K input band)—for complex agents and Chinese workloads.

ERNIE-5.0

ERNIE 5.0 native omni-modal flagship; ¥6/¥24 per 1M tokens with unified text/image/audio/video modeling.

ERNIE-5.0-Thinking-Preview

ERNIE 5.0 Thinking preview with visible chain-of-thought; per-token billing for tasks needing reasoning traces.

ERNIE-5.0-Thinking-Latest

ERNIE 5.0 Thinking latest line, same price band as Preview—for complex reasoning and agent decision nodes.

ERNIE-5.0-Thinking-Exp

ERNIE 5.0 Thinking experimental variant, priced like other 5.0 Thinking lines—availability per console.

Baidu Token API Price Comparison: Plans

ERNIE-Lite-Pro-128K
Lowest list price
Input
¥0.2
Output
¥0.4
Lite Pro lightweight tier · for high-volume baseline traffic
Usage
ERNIE-Lite-Pro-128K is the lowest listed ERNIE API entry at ¥0.2/¥0.4 per 1M tokens—for support chat, simple extraction, and high-volume light tasks.
Models
128K context with 10K RPM default rate limit—a cost-sensitive default layer; route to Turbo or 5.x when quality demands rise.
Highlights
If the main goal is budget control and throughput, Lite Pro is usually more economical than jumping straight to Turbo.
Best for
High-concurrency lightweight chat and cost-sensitive baseline traffic
ERNIE-4.5-Turbo
MainstreamRecommended
Input
¥0.8
Output
¥3.2
Cache-hit input ¥0.2 · search add-on ¥0.004/call
Usage
ERNIE-4.5-Turbo family (128K/32K/20260402 etc.) at ¥0.8/¥3.2 per 1M tokens, cache-hit input ¥0.2—a strong default for most production traffic.
Models
Supports search enhancement (¥0.004/call when triggered) and batch discounts—suited to knowledge assistants, workflow bots, and standard agents.
Highlights
If you want balance between cost and capability, Turbo is usually the most natural production default.
Best for
Teams running production chat, knowledge assistants, and cost-sensitive mainstream workloads
ERNIE-5.1
Latest flagship
Input
¥4
Output
¥18
≤32K input band · higher rates above 32K input
Usage
ERNIE-5.1 is the latest Qianfan flagship at ¥4/¥18 per 1M tokens (≤32K input band)—for complex agents, long-document understanding, and critical business nodes.
Models
128K context with native omni-modal capability—newer than ERNIE 5.0 with lower input cost in the same band, a strong flagship choice for new projects.
Highlights
Better as a key-path escalation model than routing all traffic to flagship.
Best for
Complex agents, critical business flows, and high-value output scenarios
ERNIE-5.0
Omni-modal flagship
Input
¥6
Output
¥24
Thinking variants same band · 128K context
Usage
ERNIE-5.0 and Thinking lines at ¥6/¥24 per 1M tokens (≤32K input)—native omni-modal modeling for complex reasoning and high-value tasks.
Models
Thinking Preview/Latest/Exp emit chain-of-thought—suited to agent decision stages needing visible deep reasoning.
Highlights
New deployments may prefer ERNIE-5.1; 5.0 remains for existing integrations or specific Thinking variants.
Best for
Teams running complex reasoning, omni-modal tasks, and core content generation

Baidu Token API Price Comparison: Notes

  • Official pricing shows per-thousand-token rates; this site normalizes to per-million-token units. ERNIE 5.0/5.1 charge higher bands above 32K input.
  • Turbo supports cache-hit input at ¥0.2/1M tokens plus search add-ons; some models offer discounted batch inference rates.
  • Image generation and OCR bill per image or per token—not the same unit as text API—estimate separately.
  • Beyond postpaid per-token billing, Qianfan sells prepaid volume packs for Baidu-owned lines including Lite Pro, Speed Pro, Turbo 32K/128K, Turbo VL 32K—typically 100M–5B token sizes, 6–12 month validity, with discounts (e.g. Lite Pro 100M at ¥22.5/12 mo, Turbo 128K 100M at ¥126/6 mo). Steady traffic often beats pure postpaid; this page’s entryPrice and tiers stay on postpaid unit rates—confirm pack specs on the console order page.
  • Web search also has prepaid call packs: 10,000 calls/6 mo at ¥38, 50,000 at ¥190 after discount—search enhancement draws packs first, then ¥0.004/call postpaid. If the account is overdue but token or TPM prepaid remains, inference continues and search still bills per trigger.

Baidu Token API Price Comparison: Tools & Integration

OpenAI-compatible APIAnthropic-compatible APIWeb SearchBatch inferenceContext cache

Page published

Baidu Token API Pricing, Top-up, Deals & FAQ

AI Token Plan

Global AI Token Comparison

AI Token Plan is not just a collection of vendor links. It places 41+ domestic and international AI platforms into one comparison framework, covering 395+ model entries plus common plan types such as Token Plans, Coding Plans, IDE Tools and LLM APIs.

When you need to compare AI subscription pricing, coding plan quota, API usage cost or official deal entry points, AI Token Plan brings official prices, plan tiers, usage rules, model capabilities and tool integrations into one place, reducing the need to check multiple vendor sites manually.

Data is continuously organized as vendor pricing pages, plan pages and product documentation change, making the homepage a pricing comparison entry point while detail pages explain whether each vendor plan fits individual developers, team purchasing or long-term API usage.

Prices come from each platform's official site and may change at any time; the official price prevails.

© 2026 AI Token Plan · All rights reserved · First published June 18, 2026 · 64 days running · Sitemap