Providers cap usage in dollars, credits, tokens, or messages. Pick a model and a request shape and the table converts every plan into two numbers: requests per $1, and burst — the requests per $1 you get in a single 5-hour window. Higher is better.
At GLM-5.3’s typical shape
99req / $1
on OpenCode Go · 41 plans priced · exact
41 / 41 plans
Plan details
Provider and subscription tier. Sorts by provider name.
Standard monthly price in USD, before first-month promotions.
Estimated cost per request for the selected model and token shape.
Requests fitting in the rolling 5-hour budget. '—' means no published 5-hour limit.
Requests fitting in the rolling 7-day budget. '—' means no published weekly limit.
Requests fitting in 30 days. Weekly budgets are extrapolated; unused credits do not roll over.
Requests per $1 of plan price over 30 days. Higher is better.
Requests per $1 available in one 5-hour burst. Uncapped plans match Req / $1.
SubscribeOpen provider pricing. Referral links may support TokenPlans at no extra cost.
StepFunFlash Max40,000M credits / month (no time-window or request-frequency caps)
60M input tokens/week · 2x on GLM 5/5.1/5.3, Kimi K2.7 Code, DeepSeek V4 Pro/Pro 0813 · 100 images/day
Reachability
Any agent · OpenAI-compatible key
Pricing notes
Not the best fit for coding agents — large prompts/context burn the weekly cap fast; PAYG safer for heavy coding. Web search is not included in PRO’s subscription coverage.
Cost rule
~proxy estimate · nano-gpt.com/subscription — PRO included model, fair-use cap
The $10 Go plan’s higher tier: identical rates, per-model monthly caps $60–$240 (each model’s 5h window = 20% and weekly = 50% of its cap; the headline figure is the top of that range): V4.1 Flash $120, V4 Flash $120, $60 for V4 Pro, V4 Flash Vision, Grok 4.7/4.6, Kimi K3, Qwen3.8 Max, MiMo V2.6/V2.5 Pro, GPT-6/5.6 Luna and Claude Haiku 5.5 · retention 0d (Grok 4.7, Grok 4.6, Luna 30d)
Cost rule
exact · borrowed cap · opencode.ai/docs/go — GLM-5.3 row ($1.40/$4.40/$0.26 cached read, read 2026-09-24) - Go Plus monthly limit $120 (read 2026-09-28)
Usage caps are per model $15–$60/mo (each model’s 5h window = 20% and weekly = 50% of its cap): V4.1 Flash $60, V4 Flash $30, $15 for V4 Pro, V4 Flash Vision, Grok 4.7/4.6, GLM-5.3, Kimi K3, Qwen3.8 Max, MiMo V2.6/V2.5 Pro, GPT-6/5.6 Luna and Claude Haiku 5.5 · retention 0d (Grok 4.7, Grok 4.6, Luna 30d)
Basic + premium models, covering the mainstream flagships. Non-production use (personal dev / prototyping — PAYG is the production tier). Flow ≈ $0.03283 (floats live); per-model Flow burn and per-tier model lists are not published in text. Provider "Worth" claim: $180/mo equivalent API value (≈1.8× the fee), from its published equivalent-pay-as-you-go table.
All models (Max and Ultra currently cover nearly identical lists — the difference is quota, not model selection). Non-production use (personal dev / prototyping — PAYG is the production tier). Flow ≈ $0.03283 (floats live); per-model Flow burn and per-tier model lists are not published in text. Provider "Worth" claim: $480/mo equivalent API value (≈2.4× the fee), from its published equivalent-pay-as-you-go table.
Basic models plus rotating limited-time premium models. Non-production use (personal dev / prototyping — PAYG is the production tier). Flow ≈ $0.03283 (floats live); per-model Flow burn and per-tier model lists are not published in text. Provider "Worth" claim: $30/mo equivalent API value (≈1.5× the fee), from its published equivalent-pay-as-you-go table.
$70 in credits/mo · + Smart Compression, Analytics Pro, Privacy Pack, 10 extra API keys, Custom Router, Fusion
Reachability
Any agent · OpenAI-compatible key
Pricing notes
Includes Custom Router, so you can pin a lineup and Ozore routes only among your picks. Cached-read pricing is not published, so the calculator estimates it at 10% of the input rate.
Cost rule
~proxy estimate · floor: cached reads bill at the input rate · ozore.com/pricing (live 2026-09-17 hero row, $0.91/$2.86 per 1M = 35% off maker list); cached-read not published → cached tokens billed at the input rate
OzoreBasic$20 in credits/mo · spend on any model at discounted rates · cancel anytime
$20 in credits/mo · spend on any model at discounted rates · cancel anytime
Reachability
Any agent · OpenAI-compatible key
Pricing notes
Requests are auto-routed — you do not pick the model, so a per-model figure is an upper bound on this tier; Direct Pin and Custom Router are Pro features. Cached-read pricing is not published, so the calculator estimates it at 10% of the input rate.
Cost rule
~proxy estimate · floor: cached reads bill at the input rate · ozore.com/pricing (live 2026-09-17 hero row, $0.91/$2.86 per 1M = 35% off maker list); cached-read not published → cached tokens billed at the input rate
Bills $3.99/mo after the free month · cancel anytime · entry tier (Flash-tier models) · one free trial month per person, duplicate trial accounts will be canceled · usage docs still scope a Taster key to the three floor models (deepseek-v4.1-flash, deepseek-v4-flash-0731, glm-5.3-flash); the tier card names these four
Codex: GPT-6 Luna at Standard speed in the desktop app (subject to rollout) — Go does not get GPT-6 Sol · Third-party apps use the plan through Sign in with ChatGPT (per-app weekly cap; no API key)
Price confirmed on the provider product page 2026-09-25 (dev.meta.ai/products/muse-code) · High Usage: 5× Everyday usage, more prompts with the latest Muse models
Claude Fable 5.1Claude Fable 5Claude Sonnet 5Claude Opus 5.5
Ceiling
≥5× Free · 5h rolling window + weekly cap
Reachability
—
Pricing notes
Claude Fable 5 / Fable 5.1 are usage-credits-only on Pro (billed at API rates on top of the plan) — Fable is included usage on Max 5x / Max 20x at 50% of the weekly limit
GPT-6 AstraGPT-6.1 SolGPT-6 SolGPT-6 LunaGPT-5.6 SolGPT-5.5GPT-5.6 TerraGPT-5.6 Luna
Ceiling
Unlimited* messages (reasonable use)
Reachability
—
Pricing notes
Codex usage estimates (Plus column, local messages per 5h): GPT-6 Luna 350–3,000 = highest · GPT-6.1 Sol 15–160 · GPT-6 Sol 15–150 · Astra 5–45 · weekly limits may also apply · Luna Reserve = extra Luna-only usage once the regular weekly cap is exhausted (selected Plus accounts, own limit) · Third-party apps use the plan through Sign in with ChatGPT (per-app weekly cap; no API key)
Cost rule
n/a
MiniMaxGo50% off 1st mo$11 first month, then full price (monthly billing, ends Oct 14)Usage 1× (Go baseline) · M3.1 Flash Preview included · image + audio models
Grok Bot access is included (not on Lite): link the Grok or X account to grant weekly Grok Bot usage on a Cursor account — cursor.com/help/grok-bot/plans
Cost rule
n/a
GitHub CopilotPro+$70/mo total AI credits · 4× Pro usage
Price confirmed on the provider product page 2026-09-25 (dev.meta.ai/products/muse-code) · Power Usage: 20× Everyday usage, early access to new features
Cost rule
n/a
MiniMaxExplore50% off 1st mo$27.50 first month, then full price (monthly billing, ends Oct 14)Usage 3× Go · M3.1 Flash Preview + H3 video included · image + audio models
Claude Fable 5.1Claude Fable 5Claude Opus 5.5Claude Opus 4.8
Ceiling
5× Pro · Claude Fable 5 / Fable 5.1 capped at 50% of the weekly limit
Reachability
—
Pricing notes
Claude Fable 5 / Fable 5.1 are included on Max 5x but draw only 50% of the weekly limit (claude.com/pricing “Models and usage” matrix, read 2026-10-06)
Cost rule
n/a
GitHub CopilotMax$200/mo total AI credits · 2.9× Pro+ usage
GPT-6 AstraGPT-6.1 SolGPT-6 SolGPT-6 LunaGPT-5.6 Sol ProGPT-5.6 SolGPT-5.6 TerraGPT-5.6 Luna
Ceiling
5× Plus
Reachability
—
Pricing notes
Pro plans have no five-hour limit; OpenAI publishes local-message estimates for Plus and Standard Business only · local messages and cloud chats share the allowance, and weekly limits may also apply · Third-party apps use the plan through Sign in with ChatGPT (per-app weekly cap; no API key)
MiniMaxBuild50% off 1st mo$66 first month, then full price (monthly billing, ends Oct 14)Usage 7.5× Go · M3.1 Flash Preview + H3 video included · image + audio models
Claude Fable 5.1Claude Fable 5Claude Opus 5.5Claude Opus 4.8
Ceiling
20× Pro · Claude Fable 5 / Fable 5.1 capped at 50% of the weekly limit
Reachability
—
Pricing notes
Claude Fable 5 / Fable 5.1 are included on Max 20x but draw only 50% of the weekly limit (claude.com/pricing “Models and usage” matrix, read 2026-10-06)
GPT-6 AstraGPT-6.1 SolGPT-6 SolGPT-6 LunaGPT-5.6 Sol ProGPT-5.6 SolGPT-5.6 TerraGPT-5.6 Luna
Ceiling
20× Plus
Reachability
—
Pricing notes
New Pro 200 subscriptions that are not grandfathered get a lower included usage allowance than before (grandfathered allowances run to 2026-10-29; the $200 price is unchanged) · Pro plans have no five-hour limit; OpenAI publishes local-message estimates for Plus and Standard Business only · local messages and cloud chats share the allowance, and weekly limits may also apply · Third-party apps use the plan through Sign in with ChatGPT (per-app weekly cap; no API key)
Grok Bot access is included; linking SuperGrok Heavy grants the highest linked weekly Grok Bot usage on a Cursor account (cursor.com/help/grok-bot/supergrok) — a usage grant, not a Cursor plan, and no extra usage on top of one.
Cost rule
n/a
OpenAIPro 500Highest included usage of the three Pro tiers
GPT-6 AstraGPT-6.1 SolGPT-6 SolGPT-6 LunaGPT-5.6 Sol ProGPT-5.6 SolGPT-5.6 TerraGPT-5.6 Luna
Ceiling
Highest included usage of the three Pro tiers
Reachability
—
Pricing notes
Astra Ultrafast is included only on Pro 500 among the Pro tiers ($100 / $200 / $500) — it draws the included usage allowance first, then credits, and buying credits on Pro 100 or Pro 200 does not unlock it · Fast mode consumes the included allowance at 2.5× and Astra Ultrafast at 8× (credits: 2× / 6×) · Third-party apps use the plan through Sign in with ChatGPT (per-app weekly cap; no API key)
Cost rule
n/a
No plans in the database list this model.
Some Subscribe buttons carry referral or affiliate codes — using them may give you a signup bonus or support this project at no extra cost to you.
Markers: ~ proxy estimate · ≈ coarse estimate · ≥ cache floor · ≤ cache ceiling — a plan that publishes no cached-read rate bills cached tokens at the input rate, so its Req/$1 is a floor (≥) and its cost is a ceiling (≤). Borrowed cap means the budget is a published API cap at list price, so Req / $1 is a value comparison rather than a measured allowance.
Plan price vs pay-as-you-go
Markers in these tables: ~ proxy estimate · ≈ coarse estimate · ≥ cache floor · ≤ cache ceiling — cost cells of a plan that publishes no cached-read rate are ceilings, its usage cells are floors.
+Xiaomi — subscription vs pay-as-you-go API
Xiaomi publishes both a pay-as-you-go API and four subscription tiers. The same model (MiMo V2.5) costs different amounts per token depending on how you buy it. The plan-vs-API gap explains the cross-provider "Best" pills above. Figures are at MiMo V2.5's typical shape (830 fresh + 71,500 cached + 295 output tokens). Token Plan pricing, API pricing.
Plan
Price
MiMo V2.5 $/req (amortized)
vs API rate
API-equivalent
Xiaomi API (overseas)
—
$0.000399
100% (baseline)
—
Xiaomi Lite
$6
$0.0004171
105% (5% over)
$5.74 of API
Xiaomi Standard
$16
$0.0004119
103% (3% over)
$15.50 of API
Xiaomi Pro
$50
$0.0003737
94% (6% off)
$53.38 of API
Xiaomi Max
$100
$0.0003476
87% (13% off)
$114.80 of API
OpenCode Go / ClinePass
$10 / $9.99
$0.0000665 / $0.00006643
17% of API cost
$60 of API included
At list price, Xiaomi's own tiers charge 87–105% of the Xiaomi API rate for MiMo V2.5 — Lite and Standard actually cost more than raw API; only Pro (−6%) and Max (−13%) beat it. OpenCode Go and ClinePass resell the same model at ~17% of API cost (~$60 of usage for ~$10), which is why they top the Req / $1 column. The $5–$88 signup prices are a first-month promo that renews at list.
+Z.AI — GLM Coding Plan vs OpenCode Go
Z.AI sells GLM-5.3 two ways: the GLM Coding Plan (a flat weekly credit budget) and its API, which OpenCode Go resells at the API list price ($1.40 / $4.40 / $0.26 per M tokens). The plan's credits can be valued against that API rate, which is what the numbers below do — the per-request costs are amortized at GLM-5.3's typical shape (700 fresh + 52,000 cached + 150 output tokens). Z.AI GLM Coding Plan docs.
Plan
Price
GLM-5.3 $/req (amortized)
vs API rate
API-equivalent
GLM-5.3 API (via Go)
—
$0.01516
100% (baseline)
—
GLM Coding Lite
$18
$0.004071
27% of API cost
~$67 of API
GLM Coding Pro
$80
$0.003013
20% of API cost
~$403 of API
GLM Coding Max
$168
$0.002711
18% of API cost
~$939 of API
OpenCode Go
$10
$0.01011
67% of API cost
$15 of API (published cap)
Go's flat $10 buys ~$15 of GLM-5.3 usage (2× headroom) versus ~$67–$939 for the GLM Coding Plan's $18–$168 — Go's GLM-5.3 monthly cap is just $15 of usage (GLM-5.2 keeps a $60 cap on Go), so the Coding tiers lead per dollar on the new flagship. Two caveats: credits run on a rolling weekly window with a 5-hour burst cap and never roll over, and off-peak hours (Mon–Fri 14:00–18:00 SGT) bill at 50% — roughly doubling requests then — which isn't modeled.
+Nous — Portal grants vs the OpenRouter list
Nous Portal resells OpenRouter models at 0.8× the OpenRouter list price (verified against the current list on /info), billed against a monthly credit grant with a +10% bonus: $20 → $22, $100 → $110, $200 → $220. The exception is GPT-5.6 Luna, where a limited-time -50% promo puts the Portal at half the post-repricing $0.20/$1.20 list (verified 2026-08-08). The example below amortizes each plan's subscription over its grant at GPT-5.6 Luna's typical shape (1,000 fresh + 50,000 cached + 220 output tokens) — a shape where OpenCode Go happens to charge 1× the list. Manage subscription, /info prices, API docs.
Plan
Price
GPT-5.6 Luna $/req (amortized)
vs list rate
API-equivalent
OpenRouter list (baseline)
—
$0.001464
100% (baseline)
—
Nous Portal Plus
$20
≤$0.009515
650% of list
≥~$3.08 of list usage
Nous Portal Super
$100
≤$0.009513
650% of list
≥~$15 of list usage
Nous Portal Ultra
$200
≤$0.009513
650% of list
≥~$31 of list usage
OpenCode Go
$10
$0.0009761
67% of list
~$15 of list usage ($15 cap at 1× list)
Nous Portal resells OpenRouter models at 0.8× list against a monthly grant with a +10% bonus. On most models that roughly cancels out the resellers' bundles — but wherever Go charges above list (GPT-5.6 Luna at 1×), Portal wins outright (~≥100 vs ~1,000 req/$). Caveats: unused credits roll over ($10–$100 by tier), which the monthly math ignores. A figure marked ≥ is a cache floor: the plan publishes no cached-read rate, so cached tokens bill at the input rate and the real figure is that number or higher. A cost marked ≤ is a cache ceiling: the real cost is that number or lower.
+QwenCloud — Token Plan (proxy estimate)
QwenCloud publishes two distinct prices for the same models: a PAYG rate and a Token Plan rate where credits are deducted from a single monthly quota on the plan's 30-day subscription cycle (Lite 11,500 · Essential 25,500 · Standard 45,000 · Pro 180,000 credits per cycle — the provider removed the earlier 5-hour and 7-day windows on 2026-09-22, so the table's 5h and Week columns render "—"). The docs only say the Token Plan uses "tiered deduction coefficients by model" and refer you to the subscriber console for the actual numbers — there is no public per-model credit table — so the calculator uses the published PAYG rates as a proxy. The table below shows the worked example for Qwen3.7-Plus on the Lite plan at the Grok shape (1,100 fresh + 71,500 cached + 220 output tokens):
Provider / plan
Price
5h
Week
Month
Req / $1
QwenCloud Token Plan Lite (proxy)
$8
—
—
1,765
~221
QwenCloud PAYG (proxy $/req)
—
—
—
—
$0.006512/req
OpenCode Go
$10
3,285
8,214
16,429
1,643
QwenCloud doesn't publish per-token credit costs for the Token Plan, so we estimate them from the public pay-as-you-go rates ($0.006512/req at this shape) — absolute Req / $1 figures (marked ~) carry a ≤2.25× scaling uncertainty (re-measured plan budgets price a credit at $0.000444–$0.000696 across the four tiers, 30–56% below the $0.001 proxy conversion), though relative comparisons hold. The monthly quota is the only published budget — no sub-monthly ceilings remain, unused credits don't roll over, and the preview-model capacity boost isn't modeled.
Newsletter
The weekly brief — plan churn, dated.
A plain weekly digest of every tracked price move, ceiling change and model-list update — pulled straight from the dated changelog ledger, with a one-line “worth switching?” note on each. Free. No spam, unsubscribe anytime.