Global AI Token Plan comparison home

Chatgpt Token API Pricing, Top-up, Deals & Official Entry

This OpenAI API page focuses on API key setup, official access and pricing comparison without mixing in ChatGPT Plus or Pro subscription quota: Playground, Responses API, Chat Completions, Codex CLI, Cursor, Cline and production services all use a separate platform.openai.com balance. GPT-5.6 Sol with 1,050,000 context bills at $5 input and $30 output per 1M tokens, GPT-5.6 Cyber at $12.50/$75, while GPT-5.6 Luna starts at $0.20/$1.20; prompt cache-hit and cache-write pricing, 50% Batch API savings, Flex/Fast mode, $10/1K Web Search, GPT-Realtime-2.1 voice, GPT-Image-2, Sora-2 video, rate limits and long-context multipliers all matter when judging whether OpenAI API fits as a product or coding-agent foundation.

Token APIPay-as-you-go · 1,050,000 context

Chatgpt Token API Latest Updates

Analysis

OpenAI API GPT-5.6 Pricing Comparison: Sol vs Terra vs Luna Costs, Top-Up, Rate Limits and Is It Worth It

OpenAI API pricing now comes down to the GPT-5.6 family: flagship Sol at $5 input / $30 output per 1M tokens, Terra at $2/$12, low-cost Luna at just $0.20/$1.20, and cybersecurity-focused Cyber at $12.50/$75 (short context only); cache-hit input drops to $0.02/M and Batch API adds roughly 50% off. This guide covers top-up, the official entry and API key setup, Codex CLI / Cursor / Cline integration, usage quotas and rate limits, then tells you which model is worth it for your workload.

Below is the complete pricing comparison for Chatgpt Token API, covering 9 plan tiers, core model capabilities, quotas, and official entry. All data is sourced from the official website to help you decide whether it’s worth it.

Last updated

Chatgpt Token API Price Comparison: Core Models

GPT-5.6 SolGPT-5.6 TerraGPT-5.6 LunaGPT-5.6 CyberGPT-Image-2GPT-Realtime-2.1GPT-Realtime-2.1 MiniGPT-Realtime-TranslateGPT-Realtime-WhisperGPT-Live-TranscribeSora-2
GPT-5.6 Sol

API flagship (gpt-5.6-sol, gpt-5.6 alias): Standard $5 input / $30 output per million tokens, 1,050,000 context and up to 128K output, xhigh reasoning and full agent tools—the default for complex coding and professional tasks.

GPT-5.6 Terra

More affordable professional model (gpt-5.6-terra): $2 input / $12 output per million, 1,050,000 context and up to 128K output—a balanced production workhorse.

GPT-5.6 Luna

Low-cost model (gpt-5.6-luna): $0.20 input / $1.20 output per million, 1,050,000 context and up to 128K output—low latency and cost for sub-agents and light volume.

GPT-Image-2

Latest image generation/editing API billed by image and text modalities (e.g. $30/M image output)—used with images/generations and images/edits endpoints.

GPT-5.6 Cyber

Daybreak cybersecurity model (gpt-5.6-cyber): $12.50 input / $75 output per million, short context only, for vulnerability analysis, threat detection, and security agents.

Chatgpt Token API Price Comparison: Plans

GPT-5.6 Sol
Flagship reasoningRecommended
Input
$5
Output
$30
Input $5/M · output $30/M · cache-hit input $0.50/M · 1,050,000 context
Usage
Model id gpt-5.6-sol (gpt-5.6 alias)—the current API flagship at $5/M input and $30/M output (Standard), 1,050,000 context, up to 128K output—suited to complex reasoning, professional coding, and multi-step agents.
Models
Supports reasoning.effort (none/low/medium/high/xhigh), Functions, Web search, File search, Computer use, Code interpreter, and more via Chat Completions or the Responses API.
Highlights
Cache-hit input is only $0.50/M; sessions with >272K input tokens bill at higher multipliers—enable prompt caching for repeated prefixes.
Best for production apps treating OpenAI as a core base—Codex-style coding agents, long-document analysis, and high-stakes knowledge workflows.
Best for
Complex reasoning and coding, production agent systems, and developers needing the strongest API tier
GPT-5.6 Terra
Balanced flagship
Input
$2
Output
$12
Input $2/M · output $12/M · cache-hit input $0.20/M · 1,050,000 context
Usage
Model id gpt-5.6-terra—a more affordable professional workhorse at $2/M input and $12/M output with 1,050,000 context and up to 128K output, sitting between Sol and Luna on capability and cost.
Models
Also supports Functions, Web search, File search, Computer use, and similar tools—good for high-volume production where you need strong capability at lower unit cost.
Highlights
Cache-hit input is $0.20/M; Batch API adds ~50% savings—suited to offline batch and delay-tolerant jobs.
When GPT-5.6 Sol unit cost is too high but Luna is not enough, Terra is usually the API sweet spot.
Best for
Mid-complexity production calls needing 1,050,000 context at lower cost than Sol
GPT-5.6 Luna
Fast & low-cost
Input
$0.20
Output
$1.20
Input $0.20/M · output $1.20/M · cache-hit input $0.02/M · 1,050,000 context
Usage
Model id gpt-5.6-luna—the low-cost tier at $0.20/M input and $1.20/M output, 1,050,000 context and up to 128K output—for coding, computer use, and sub-agents.
Models
Lower latency and unit cost—suited to high-frequency light completion, routing/classification, batch formatting, and cost-sensitive volume.
Highlights
Cache-hit input is as low as $0.02/M; choose Luna when optimizing latency and cost.
Good for sub-agents, intermediate pipeline steps, and tiered architectures that start on Luna and escalate hard tasks to Sol.
Best for
High-volume batch calls, sub-agents, cost-sensitive production completion, and light reasoning
GPT-5.6 Cyber
Cybersecurity expert
Input
$12.50
Output
$75
Input $12.50/M · output $75/M · cache-hit input $1.25/M · cache write $15.625/M · short context only
Usage
Model id gpt-5.6-cyber—Daybreak series cybersecurity model at $12.50/M input and $75/M output, short context only (≤270K), for vulnerability analysis, threat detection, and security agents.
Models
Cache-hit input is $1.25/M, cache writes $15.625/M—suited to high-frequency repeat calls with the same security context.
Highlights
Specialized for security vs GPT-5.6 Sol—higher unit cost but domain-focused; choose Sol for general coding and reasoning.
Best for
Security research, vulnerability analysis, threat intelligence, and compliance audit teams
GPT-Image-2
Image generation
Input
$8
Output
$30
Image input $8/M · image output $30/M · text input $5/M · billed per modality
Usage
GPT-Image-2 is the latest image generation and editing model—image modality at $8/M input and $30/M output, text input at $5/M (each with cache-hit rates).
Models
Called via v1/images/generations and v1/images/edits—suited to in-app image generation and multimodal workflows; do not assume pure chat input/output rates.
Highlights
Images are tokenized for billing—per-image cost depends on resolution and prompt complexity; estimate with Playground or small tests before launch.
Best for
Products and creative workflows needing official image generation/editing APIs
GPT-Realtime-2.1
Realtime voice
Input
$4
Output
$24
Text $4/$24 per M · audio $32/$64 per M · image $5/M input · cache-hit $0.40/M · multi-modality billing
Usage
GPT-Realtime-2.1 targets realtime voice—text at $4/M input and $24/M output; audio at $32/M input and $64/M output; image input at $5/M; audio and text cache-hit at $0.40/M.
Models
Integrates via v1/realtime sessions—for voice assistants, support bots, and low-latency conversational products; cost depends on audio duration and text mix.
Highlights
Billing differs from pure text chat—compare total cost against a Transcribe + chat + TTS pipeline. Also available: GPT-Realtime-2.1 Mini at lower cost (text $0.60/$2.40, audio $10/$20).
Best for
Realtime voice products, low-latency dialogue, and multimodal voice assistant developers
GPT-Realtime-2.1 Mini
Lightweight realtime
Input
$0.60
Output
$2.40
Text $0.60/$2.40 per M · audio $10/$20 per M · image $0.80/M input · cache-hit $0.06–$0.30/M
Usage
GPT-Realtime-2.1 Mini is the lower-cost alternative to 2.1: text at $0.60/M input and $2.40/M output, audio at $10/M input and $20/M output, image input at $0.80/M.
Models
Suited to lightweight voice interaction, live support, and budget voice agents—balancing latency and cost.
Best for
Budget realtime voice apps, lightweight support, and voice agent prototypes
GPT-Realtime-Translate
Live translation
Price Comparison
$0.034
$0.034/min ($0.00057/sec) · live speech translation
Usage
GPT-Realtime-Translate offers live speech translation at $0.034/min ($0.00057/sec)—suited to meetings, streaming, and multilingual support.
Models
Billed by audio duration rather than input/output token rates—estimate minute-level cost for long-running sessions.
Best for
Live interpretation, cross-language meetings, and streaming translation products
GPT-Realtime-Whisper
Streaming STT
Price Comparison
$0.017
$0.017/min ($0.00028/sec) · streaming transcription
Usage
GPT-Realtime-Whisper streams speech to text at $0.017/min ($0.00028/sec) as the speaker talks.
Models
Suited to live captions, meeting notes, and voice input pipelines—different from batch models like GPT-4o Transcribe; pick by latency needs.
Best for
Live captions, meeting dictation, and low-latency speech input apps

Chatgpt Token API Price Comparison: Notes

  • Prices below are Standard processing for context under 270K; Batch API is ~50% off input/output; data residency adds 10% for the GPT-5.6 family. GPT-5.6 prompts over 272K input tokens bill 2x input and 1.5x output for the full request. Cache writes are billed separately: GPT-5.6 Sol $6.25/M, Terra $2.50/M, Luna $0.25/M, Cyber $15.625/M. Priority processing was renamed to Fast mode on 2026-07-30.
  • Prompt cache-hit input: GPT-5.6 Sol $0.50/M, GPT-5.6 Terra $0.20/M, GPT-5.6 Luna $0.02/M—design caching for repeated system prompts and long prefixes.
  • Web Search tool is $10 per 1k calls (search content tokens free); Containers bill by container size (switching to 20-minute sessions from 2026-03-31).
  • ChatGPT Plus/Business/Enterprise subscriptions do not include standard API usage; Playground and production API share balance and bill per dashboard usage reports.

Chatgpt Token API Price Comparison: Tools & Integration

OpenAI APIResponses APIChat CompletionsCodex CLICursorCline

Chatgpt Token API Pricing, Top-up, Deals, Quota, Usage, Setup & Updates

Page published

Chatgpt Token API Pricing, Top-up, Deals & FAQ

AI Token Plan

Global AI Token Comparison

AI Token Plan is not just a collection of vendor links. It places 41+ domestic and international AI platforms into one comparison framework, covering 395+ model entries plus common plan types such as Token Plans, Coding Plans, IDE Tools and LLM APIs.

When you need to compare AI subscription pricing, coding plan quota, API usage cost or official deal entry points, AI Token Plan brings official prices, plan tiers, usage rules, model capabilities and tool integrations into one place, reducing the need to check multiple vendor sites manually.

Data is continuously organized as vendor pricing pages, plan pages and product documentation change, making the homepage a pricing comparison entry point while detail pages explain whether each vendor plan fits individual developers, team purchasing or long-term API usage.

Prices come from each platform's official site and may change at any time; the official price prevails.

© 2026 AI Token Plan · All rights reserved · First published June 18, 2026 · 64 days running · Sitemap