Openrouter Token Plan Pricing, Top-up, Deals & Updates | Article

Openrouter Token Plan subscription pricing. 15 tiers (Credits Prepaid, from Free/mo). with Prepaid Credits. GPT-5.6 Sol and 64 more models. OpenRouter Chat, REST API, OpenAI SDK and 2 more tools integrations. Plan quotas, perks, and official entry—is it worth it?

41
Platforms
395+
Models
140+
IDE Tools
65
Openrouter Token Plan models
Free
Openrouter Token Plan entry price
View
Vendor details

OpenRouter Credits Latest Updates

OpenRouter Pricing Comparison: Which of 400+ Models Is Cheapest? Per-Provider Rates, 5.5% Top-Up Fee, and Is It Worth It

Published Updated
AnalysisOpenRouterPricing ComparisonModel PricingCreditsTop-Up FeeGPTClaudeDeepSeek
Summary

OpenRouter has no monthly plan—only prepaid Credits with pay-as-you-go billing. Its 400+ models bill at provider pass-through rates with no markup, but top-ups carry a 5.5% fee ($0.80 minimum). The cheapest paid model, Qwen Flash, starts at $0.03/1M input tokens, while GPT-5.6 Sol tops out at $30/1M output—nearly a 50x gap between flagships. This article tallies entry and flagship prices across 15 providers, free-model limits (50→1000/day), and the direct-API vs OpenRouter math so you can decide whether it's worth it.

OpenRouter pricing comparison isn't about "which plan is cheaper" — it's about three things: which of 400+ models costs less, whether the platform marks up prices, and what top-ups cost. There's no monthly subscription, only prepaid Credits billed per token at pass-through rates with no markup, but top-ups carry a 5.5% fee ($0.80 minimum). The cheapest paid model, Qwen Flash, starts at $0.03/1M input tokens, while GPT-5.6 Sol costs $30/1M output — nearly a 50x spread. This article tallies entry and flagship prices across 15 providers, free-model limits, and the direct-API vs OpenRouter math. Prices below are as of 2026-08-13; check OpenRouter for live rates.

Billing model: no monthly plan, just prepaid Credits

OpenRouter has no monthly or annual subscription, only prepaid Credits. After you top up at settings/credits, the API and OpenRouter Chat share one USD balance billed per model at its token rate (source). A real price comparison has to look at three parts: model token rates, the top-up fee, and free-model quota.

Billing item Rule Notes
Inference token rate Matches provider list price No markup
Top-up fee 5.5% ($0.80 minimum) 5% for crypto
Free-model quota 50/day by default 1000/day after $10 in credits
BYOK service fee $25,000 list-price inference/month free ($200,000 Enterprise) 5% beyond that
Logging discount 1% usage discount for optional prompt/completion logging Optional

15-provider price comparison: 100x entry gap, nearly 50x flagship gap

The cheapest paid model on OpenRouter is Qwen Flash at $0.03/1M input, and the priciest flagship is GPT-5.6 Sol at $30/1M output — more than an order of magnitude apart. Here are the 15 core providers, all priced per 1M tokens (input/output) (source):

Provider Entry tier (input/output) Flagship tier (input/output)
OpenAI Luna $0.10/$0.60 Sol $5/$30
Anthropic Haiku 4.5 $1/$5 Opus 5 $5/$25
Google Flash Lite $0.30/$2.50 3.6 Flash $1.5/$7.5
DeepSeek Flash 0731 $0.08/$0.18 Pro 0813 $0.435/$0.87
xAI Build $1/$2 Grok 4.6 $2/$6
Qwen 3.7 Flash $0.03/$0.13 3.8 Max $2/$6
Z.ai GLM 4.7 Flash $0.06/$0.40 5.2 $0.49/$1.54
MiniMax M2.5 $0.22/$0.90 M3 $0.30/$1.20
Kimi K2.5 $0.57/$2.85 K3 $3/$15
Xiaomi MiMo V2.5 $0.14/$0.28 V2.5-Pro $0.435/$0.87
Meta Scout $0.10/$0.30 Maverick $0.20/$0.70
Mistral Small 4 $0.15/$0.60 Medium 3.5 $1.5/$7.5
Perplexity Sonar $1/$1 Pro Search $3/$15
NVIDIA Nano $0.05/$0.20 Ultra $0.60/$3.60
Tencent Hy3 preview $0.06/$0.21 Hy3 $0.132/$0.528

Two patterns stand out. Chinese providers keep entry prices very lowDeepSeek, Qwen, Z.ai GLM, MiMo, and Tencent Hunyuan all price input under $0.10/M, making them ideal for high volume. Flagships, meanwhile, get expensive — Claude Opus 5 costs $5/$25 and Kimi K3 $3/$15, 20-30x their own entry tiers.

How to pick: match the model to the job

Don't just chase the cheapest option — OpenRouter's real value is switching between 400+ models with one API key, so "which one to pick" matters more than "which is cheapest." Break it into three scenarios:

High volume / concurrency: prefer low-cost Chinese models

For large-scale inference, batch processing, and agent subtasks, DeepSeek Flash at $0.08/$0.18, Qwen 3.7 Flash at $0.03/$0.13, and Z.ai GLM 4.7 Flash at $0.06/$0.40 are outstanding value. They rank among the most-used models on OpenRouter and deliver far lower long-context cost than peer closed flagships.

Coding and agents: Claude, GPT, Grok, Kimi

For writing code and long-horizon agents, Claude (Opus 5 $5/$25, Sonnet 5 $2/$10) and OpenAI's GPT-5.3-Codex ($1.75/$14) are the defaults. Grok Build 0.1 is a $1/$2 coding agent, while Kimi K3 offers ~1M context at $3/$15 for ultra-long documents and UI/code generation.

Multimodal and search: Gemini, Perplexity, MiniMax

For image/audio/video input, Gemini 3.x starts at Flash Lite $0.30/$2.50. For web search and cited Q&A, Perplexity Sonar starts at $1/$1. For multimodal agents, look at MiniMax M3 at $0.30/$1.20.

Direct API vs OpenRouter: is the 5.5% fee worth it?

OpenRouter doesn't profit from token spread — it only charges the 5.5% top-up fee. If you switch between many models, that fee is cheaper than the hassle of direct accounts; if you stick to one model, direct is cheaper. Here's the math (source):

Dimension Direct vendor OpenRouter
Token rate Official price Same as official, no markup
Platform fee None 5.5% top-up fee ($0.80 min)
Model coverage Single vendor 400+ models, one balance
Switching cost Separate account and top-up per vendor One key for all
Failover None Provider fallback

The takeaway is clear: heavy single-model users should go direct to save that 5.5%; multi-model switchers, evaluators, and anyone needing provider fallback pay the fee for convenience and get their money's worth.

Caveats

  1. Credits expire: unused Credits may expire one year after purchase — don't top up too much at once.
  2. Refunds are limited: refunds within 24 hours on the Credits page, but platform fees are non-refundable and crypto is never refundable.
  3. Free models are capped: 50/day by default, 1000/day after $10 in credits — fine for testing, not production.
  4. BYOK has a service fee: your own provider keys get a monthly no-fee allowance measured by list-price inference cost ($25,000/month on pay-as-you-go, $200,000/month on Enterprise), then 5% beyond that.
  5. Rates change: all prices pass through openrouter.ai/models, so actual billing follows the live rate.

Sources and verification

Last verified

Openrouter Token Plan FAQ

Which model is cheapest in an OpenRouter pricing comparison, and which flagship is the most expensive?

The cheapest paid model is Qwen 3.7 Flash at $0.03/1M input and $0.13/1M output tokens; the most expensive flagship is OpenAI GPT-5.6 Sol at $5 input and $30 output per 1M tokens. Between them sit budget tiers like DeepSeek Flash ($0.08/$0.18) and MiMo V2.5 ($0.14/$0.28). Check openrouter.ai/models for the full 400+ model catalog.

How is OpenRouter's 5.5% top-up fee calculated, and what is the minimum charge?

OpenRouter bills inference tokens at provider pass-through rates with no markup; the platform only charges a 5.5% fee on credit purchases with a $0.80 minimum. That means a $10 top-up actually costs about $10.55 and $100 costs about $105.50. Crypto top-ups carry a 5% fee instead.

What is the difference between the 50/day and 1000/day free-model limits on OpenRouter?

Free model API access defaults to 50 requests per day; after you have purchased $10 or more in credits, the free-model limit rises to 1000 requests per day. The quota resets daily and is meant for prototyping and zero-cost model testing, not production.

Is OpenRouter cheaper than using the official GPT or Claude API directly?

If you use only one or two models with steady volume, direct provider APIs have no top-up fee and are cheaper. If you switch between many models and want a unified balance plus provider fallback, OpenRouter's 5.5% fee buys convenience that is worth it. It comes down to how often you switch models.

Do OpenRouter Credits expire, and can I get a refund after topping up?

Unused Credits may expire one year after purchase. Refunds can be requested within 24 hours on the Credits page, but platform fees are non-refundable and crypto top-ups are never refundable.
AI Token Plan

Global AI Token Comparison

AI Token Plan is not just a collection of vendor links. It places 41+ domestic and international AI platforms into one comparison framework, covering 395+ model entries plus common plan types such as Token Plans, Coding Plans, IDE Tools and LLM APIs.

When you need to compare AI subscription pricing, coding plan quota, API usage cost or official deal entry points, AI Token Plan brings official prices, plan tiers, usage rules, model capabilities and tool integrations into one place, reducing the need to check multiple vendor sites manually.

Data is continuously organized as vendor pricing pages, plan pages and product documentation change, making the homepage a pricing comparison entry point while detail pages explain whether each vendor plan fits individual developers, team purchasing or long-term API usage.

Prices come from each platform's official site and may change at any time; the official price prevails.

© 2026 AI Token Plan · All rights reserved · First published June 18, 2026 · 64 days running · Sitemap