Global AI Token Plan comparison home

Claude Token API Pricing, Top-up, Deals & Official Entry

Use this Claude API page to compare Anthropic official API key access, pricing and deals: Haiku 4.5 starts at $1 input and $5 output per 1M tokens, Sonnet 5 is now standard at $2 input and $10 output (the planned $3/$15 increase was cancelled), Opus 5 covers 1M-context enterprise agents, and Fable 5 is the next-generation flagship for long-running agents. It connects the official Console entry point, Claude Code / Cursor / Cline setup, prompt caching, 50% Batch API savings, US-only inference at 1.1x, Web Search, Web Fetch, Managed Agents, Bedrock / Vertex AI regional premiums and CCU postpaid billing so you can compare quota, rate limits and whether to apply for a Claude API key directly.

Token APIPay-as-you-go · 1M context

Claude Token API Latest Updates

Pricing

Claude Sonnet 5 Price Hike Cancelled: Stays at $2/$10, Four-Model Pricing Comparison

Anthropic has cancelled the planned September 1 increase of Claude Sonnet 5 from $2/$10 to $3/$15, so $2/$10 is now the standard price. This pricing comparison breaks down how Sonnet 5 tiers against Haiku 4.5, Opus 5 and Fable 5, how much prompt caching and Batch API can save, and benchmarks it against DeepSeek and OpenAI, together with the official entry point, API key setup and integration config so you can judge whether Sonnet 5 is worth keeping.

Below is the complete pricing comparison for Claude Token API, covering 4 plan tiers, core model capabilities, quotas, and official entry. All data is sourced from the official website to help you decide whether it’s worth it.

Last updated

Claude Token API Price Comparison: Core Models

Claude Fable 5Claude Opus 5Claude Sonnet 5Claude Haiku 4.5
Claude Fable 5

Top API tier (claude-fable-5): $10 input / $50 output per MTok, 1M context, 128K max output, adaptive thinking always on, for long-running agents.

Claude Opus 5

Complex enterprise work (claude-opus-5): $5 input / $25 output per MTok, 1M context, 128K max output, for complex agentic coding and enterprise work.

Claude Sonnet 5

High-performance production workhorse (claude-sonnet-5): standard $2 / $10 per MTok (the planned $3 / $15 increase was cancelled), 1M context, 128K max output.

Claude Haiku 4.5

Fast low-cost tier (claude-haiku-4-5): $1 / $5 per MTok, 200K context, 64K max output, for light and high-concurrency first-pass tasks.

Claude Token API Price Comparison: Plans

Claude Fable 5
Long-running agent flagshipRecommended
Input
$10
Output
$50
Input $10/M · output $50/M · cache write $12.50/M · cache hit $1/M · 1M context
Usage
Model id claude-fable-5: $10/M input and $50/M output, 1M context, up to 128K output.
Models
Adaptive thinking is always on; positioned as next-generation intelligence for long-running agents.
Highlights
5-minute cache writes $12.50/M and reads $1/M.
Suited to long task chains, complex autonomous agents, and high-value knowledge work.
Best for
Long-running agents, top-intelligence tasks, and high-value production workflows
Claude Opus 5
Complex enterprise work
Input
$5
Output
$25
Input $5/M · output $25/M · cache write $6.25/M · cache hit $0.50/M · 1M context
Usage
Model id claude-opus-5: $5/M input and $25/M output, 1M context, up to 128K output.
Models
Positioned for complex agentic coding and enterprise work, suited to complex code, enterprise workflows, and high-autonomy tasks.
Highlights
Cache writes $6.25/M and reads $0.50/M.
When Fable 5 is too costly but flagship capability is still needed, Opus 5 is the lower-cost choice.
Best for
Complex agents, long-horizon coding, and enterprise knowledge work
Claude Sonnet 5
High-performance coding and agents
Input
$2
Output
$10
Input $2/M · output $10/M · cache write $2.50/M · cache hit $0.20/M · 1M context
Usage
Model id claude-sonnet-5: standard $2/M input and $10/M output; the planned $3/$15 increase was cancelled.
Models
1M context, up to 128K output, adaptive thinking, positioned for high-performance coding and agents.
Highlights
Cache writes $2.50/M and reads $0.20/M.
Suited to high-frequency API calls, shared team backends, and 1M-context workloads that do not need Fable/Opus throughout.
Best for
Daily production API, coding agents, and team backends balancing cost and capability
Claude Haiku 4.5
Fast & low-cost
Input
$1
Output
$5
Input $1/M · output $5/M · cache write $1.25/M · cache hit $0.10/M · 200K context
Usage
Model id claude-haiku-4-5 (snapshot claude-haiku-4-5-20251001): $1/M input and $5/M output, 200K context, up to 64K output, lowest latency.
Models
Near-frontier intelligence at the lowest cost—for quick Q&A, routing/classification, formatting, sub-agents, and lightweight pipeline steps.
Highlights
Supports extended thinking (no adaptive thinking); cache writes $1.25/M, cache hits $0.10/M, with further savings via Batch API.
Suited to tiered architectures—Haiku first pass, Sonnet/Opus for hard tasks—and cost-sensitive batch workloads.
Best for
High-volume light calls, sub-agents, routing, and cost-sensitive batch tasks

Claude Token API Price Comparison: Notes

  • Standard Claude API input/output list prices (USD per 1M tokens); prompt cache writes are 1.25x input (5-minute) and 2x input (1-hour), with reads at 0.1x input.
  • Batch API is ~50% off input/output with async completion within 24 hours; Priority Tier adds higher throughput at official Priority multipliers.
  • US-only inference, where available, bills input/output at 1.1x. Bedrock/Vertex regional and multi-region endpoints may add 10%.
  • Legacy Sonnet 4.6, Opus 4.7, Opus 4.6, Sonnet 4.5, Opus 4.5, Opus 4.1 and others may still appear in pricing tables; new integrations should prefer Fable 5 / Opus 5 / Sonnet 5 / Haiku 4.5.
  • Additional pricing: Fast Mode (Opus 5 / Opus 4.8) $10 input / $50 output; Managed Agents $0.08/session-hour; Web Search $10/1K searches; Code Execution is free with Web Search / Web Fetch, otherwise 1,550 free hours/month/org then $0.05/hour/container. Prompt Caching also has an Extended mode.

Claude Token API Price Comparison: Tools & Integration

Anthropic APIMessages APIClaude CodeCursorClineAmazon BedrockVertex AIMicrosoft FoundryManaged AgentsWeb SearchWeb Fetch

Claude Token API Pricing, Top-up, Deals, Quota, Usage, Setup & Updates

Page published

Claude Token API Pricing, Top-up, Deals & FAQ

AI Token Plan

Global AI Token Comparison

AI Token Plan is not just a collection of vendor links. It places 41+ domestic and international AI platforms into one comparison framework, covering 395+ model entries plus common plan types such as Token Plans, Coding Plans, IDE Tools and LLM APIs.

When you need to compare AI subscription pricing, coding plan quota, API usage cost or official deal entry points, AI Token Plan brings official prices, plan tiers, usage rules, model capabilities and tool integrations into one place, reducing the need to check multiple vendor sites manually.

Data is continuously organized as vendor pricing pages, plan pages and product documentation change, making the homepage a pricing comparison entry point while detail pages explain whether each vendor plan fits individual developers, team purchasing or long-term API usage.

Prices come from each platform's official site and may change at any time; the official price prevails.

© 2026 AI Token Plan · All rights reserved · First published June 18, 2026 · 64 days running · Sitemap