Gemini Token API Pricing, Top-up, Deals & Official Entry
Use this Gemini API page to compare Google AI Studio / Vertex AI official access, API key setup, pricing and free quota rather than Google AI Plus / Pro / Ultra subscription limits: Gemini 3.1 Flash-Lite starts at $0.25 input and $1.50 output per 1M tokens, Gemini 3.6 Flash is the latest $1.50/$7.50 fast flagship replacing 3.5 Flash as the default recommendation, and Gemini 3.1 Pro uses ≤200K and >200K context price bands. It connects Paid tier rate limits, 50% Batch API savings, context caching storage fees, Flex / Priority, Grounding with 5,000 free prompts per month then $14/1K queries, Live API and image-generation billing so you can judge Gemini API for production apps, AI Studio testing or search-augmented agents.
Below is the complete pricing comparison for Gemini Token API, covering 9 plan tiers, core model capabilities, quotas, and official entry. All data is sourced from the official website to help you decide whether it’s worth it.
Last updated:
Gemini Token API Price Comparison: Core Models
Latest fast API flagship (gemini-3.6-flash): $1.50 input / $7.50 output per MTok, faster and cheaper than 3.5 Flash—Free tier for trials.
Previous fast flagship (gemini-3.5-flash): $1.50 input / $9 output per MTok, frontier intelligence + search grounding—prefer 3.6 Flash for new projects.
Strongest Pro preview (gemini-3.1-pro-preview): $2/$12 per MTok (≤200K), multimodal agents and vibe-coding—Paid only.
Best value (gemini-3.1-flash-lite): $0.25/$1.50 per MTok—first choice for high-volume agents and translation; consider upgrading to 3.5 Flash-Lite.
New-gen lite (gemini-3.5-flash-lite): $0.30/$2.50 per MTok—the upgrade from 3.1 Flash-Lite with improved model capability.
Gemini Token API Price Comparison: Plans
Default for search-grounded fast agent loops and daily production API—the direct upgrade from 3.5 Flash; escalate to 3.1 Pro for complex multimodal agents.
For existing 3.5 Flash integrations; new projects should use 3.6 Flash for lower output cost and higher speed.
Same model family as gemini.google subscription 3.1 Pro, but API bills per token independently of Plus/Pro/Ultra usage multipliers.
Audio input Standard $0.50/M, Batch $0.25/M—estimate separately for speech pipelines.
Existing 2.5 Pro integrations can keep billing; migration plans should weigh 3.1 Pro multimodal agent gains.
Suited to production needing controllable thinking depth and 1M context without 3.5 Flash pricing.
Gemini Token API Price Comparison: Notes
- Prices below are Paid tier Standard processing in USD per 1M tokens; Free tier in AI Studio offers free input/output on select models (content may improve products)—upgrade to Paid for production.
- Gemini 3.1 Pro and 2.5 Pro use tiered pricing at ≤200K vs >200K prompt tokens (e.g. 3.1 Pro Standard: $2/$12 vs $4/$18 per MTok).
- Batch API is ~50% off input/output; context caching adds storage fees (typically $0.50–$4.50 per 1M tokens/hour, model-dependent).
- Image generation (3.1 Flash Image, etc.) and Live API audio/video use different billing from plain text chat—estimate per use case before integration.
Gemini Token API Price Comparison: Tools & Integration
Page published: