Claude Token API Pricing, Top-up, Deals & Official Entry
Use this Claude API page to compare Anthropic official API key access, pricing and deals: Haiku 4.5 starts at $1 input and $5 output per 1M tokens, Sonnet 5 is now standard at $2 input and $10 output (the planned $3/$15 increase was cancelled), Opus 5 covers 1M-context enterprise agents, and Fable 5 is the next-generation flagship for long-running agents. It connects the official Console entry point, Claude Code / Cursor / Cline setup, prompt caching, 50% Batch API savings, US-only inference at 1.1x, Web Search, Web Fetch, Managed Agents, Bedrock / Vertex AI regional premiums and CCU postpaid billing so you can compare quota, rate limits and whether to apply for a Claude API key directly.
Below is the complete pricing comparison for Claude Token API, covering 4 plan tiers, core model capabilities, quotas, and official entry. All data is sourced from the official website to help you decide whether it’s worth it.
Last updated:
Claude Token API Price Comparison: Core Models
Top API tier (claude-fable-5): $10 input / $50 output per MTok, 1M context, 128K max output, adaptive thinking always on, for long-running agents.
Complex enterprise work (claude-opus-5): $5 input / $25 output per MTok, 1M context, 128K max output, for complex agentic coding and enterprise work.
High-performance production workhorse (claude-sonnet-5): standard $2 / $10 per MTok (the planned $3 / $15 increase was cancelled), 1M context, 128K max output.
Fast low-cost tier (claude-haiku-4-5): $1 / $5 per MTok, 200K context, 64K max output, for light and high-concurrency first-pass tasks.
Claude Token API Price Comparison: Plans
Suited to long task chains, complex autonomous agents, and high-value knowledge work.
When Fable 5 is too costly but flagship capability is still needed, Opus 5 is the lower-cost choice.
Suited to high-frequency API calls, shared team backends, and 1M-context workloads that do not need Fable/Opus throughout.
Suited to tiered architectures—Haiku first pass, Sonnet/Opus for hard tasks—and cost-sensitive batch workloads.
Claude Token API Price Comparison: Notes
- Standard Claude API input/output list prices (USD per 1M tokens); prompt cache writes are 1.25x input (5-minute) and 2x input (1-hour), with reads at 0.1x input.
- Batch API is ~50% off input/output with async completion within 24 hours; Priority Tier adds higher throughput at official Priority multipliers.
- US-only inference, where available, bills input/output at 1.1x. Bedrock/Vertex regional and multi-region endpoints may add 10%.
- Legacy Sonnet 4.6, Opus 4.7, Opus 4.6, Sonnet 4.5, Opus 4.5, Opus 4.1 and others may still appear in pricing tables; new integrations should prefer Fable 5 / Opus 5 / Sonnet 5 / Haiku 4.5.
- Additional pricing: Fast Mode (Opus 5 / Opus 4.8) $10 input / $50 output; Managed Agents $0.08/session-hour; Web Search $10/1K searches; Code Execution is free with Web Search / Web Fetch, otherwise 1,550 free hours/month/org then $0.05/hour/container. Prompt Caching also has an Extended mode.
Claude Token API Price Comparison: Tools & Integration
Claude Token API Pricing, Top-up, Deals, Quota, Usage, Setup & Updates
Page published: