All plans/Calculator/GLM-5.3-Flash

GLM-5.3-Flash Calculator — requests per $1

Every monthly plan that includes GLM-5.3-Flash, converted to requests per $1 at the model's typical request shape. Same engine as the main calculator — proxy estimates carry a ~ marker; coarse estimates carry .

GLM-5.3-Flash (formerly the stealth "Ox Alpha" codename on OpenRouter/CommandCode) is a confirmed Z.AI GLM-family model — it appears in Z.AI’s own "Latest Models" pricing (docs.z.ai/guides/overview/pricing) at $0.15 input / $0.075 cached / $0.50 output, on a 50% launch discount through 2026-09-09. 1M context, multimodal input (text, images, video), built for long-horizon coding and agentic work; free during its ~1-week CommandCode preview window on Go and above.

At GLM-5.3-Flash’s typical shape

12,048req / $1

on Ozore Pro · 14 plans priced · exact

Agent Arena benchmark

#21 · 2% net improvement

Agent Arena leaderboard · fetched 2026-09-10

Per-request shape · tokens

Throttling, off-peak, rollover, and cache writes are not modeled.

Price$
15 / 15 plans
Plan detailsProvider and subscription tier. Sorts by provider name.Standard monthly price in USD, before first-month promotions.Estimated cost per request for the selected model and token shape.Requests fitting in the rolling 5-hour budget. '—' means no published 5-hour limit.Requests fitting in the rolling 7-day budget. '—' means no published weekly limit.Requests fitting in 30 days. Weekly budgets are extrapolated; unused credits do not roll over.Requests per $1 of plan price over 30 days. Higher is better.Requests per $1 available in one 5-hour burst. Uncapped plans match Req / $1.SubscribeOpen provider pricing. Referral links may support TokenPlans at no extra cost.
OzorePro$70 in credits/mo · + Smart Compression, Analytics Pro, Privacy Pack, 10 extra API keys, Custom Router, Fusion$35.00$0.0002421,68612,04812,048Subscribe
OzoreBasic$20 in credits/mo · spend on any model at discounted rates · cancel anytime$10.00$0.0002120,48112,04812,048Subscribe
CommandCodeGoCheapest$10 credits · ~15K reqs/mo · $3/5h · $6/7d$1.00$0.00446851,3712,285~2,285~685Subscribe
CommandCodeGOATNew$70 credits · ~75K reqs/mo · $14/5h · $35/7d$10.00$0.00443,1997,99915,999~1,600~320Subscribe
Z.aiGLM Coding Max140,000 Credits / 7 days (14× Lite) · 28,000 Credits / 5 hours$168.00$0.00108,06940,345172,9071,02948Subscribe
Z.aiGLM Coding Pro60,000 Credits / 7 days (6× Lite) · 12,000 Credits / 5 hours$80.00$0.00113,45817,29174,10492643Subscribe
CommandCodePro$80 credits · ~100K reqs/mo · $16/5h · $40/7d$20.00$0.00443,6579,14218,285~914~183Subscribe
OpenCodeGo$60/mo usage · $30/wk · $12/5h$10.00$0.00191,5783,9477,894789158Subscribe
Z.aiGLM Coding Lite10,000 Credits / 7 days · 2,000 Credits / 5 hours$18.00$0.00155762,88112,34768632Subscribe
OllamaMax10 concurrent requests · $300/mo usage credits$100.00$0.001952,631526526Subscribe
OllamaPro$60/mo usage credits · access to larger pro models$20.00$0.001910,526526526Subscribe
CommandCodeMax 20×$300 credits · ~370K reqs/mo · $90/5h · $180/7d$200.00$0.004420,57141,14268,571~343~103Subscribe
CommandCodeMax 10×$150 credits · ~230K reqs/mo · $45/5h · $90/7d$100.00$0.004410,28520,57134,285~343~103Subscribe
CommandCodeProviderPAYG · top-ups roll over, never expire · zero markup$15.00$0.00443,428~229~229Subscribe
SyntheticSubscription500 requests/5h per pack · 1 concurrent request per model · +500 req/5h per $30 pack (max 5 packs)$30.00n/an/an/an/aSubscribe

Some Subscribe buttons carry referral or affiliate codes — using them may give you a signup bonus or support this project at no extra cost to you.

Markers: ~ proxy estimate and coarse estimate. Borrowed cap means the budget is a published API cap at list price, so Req / $1 is a value comparison rather than a measured allowance.

Best-priced plans

Best-priced plans for GLM-5.3-Flash

The top 5 plans by computed Req/$1 that include GLM-5.3-Flash, plus the lowest-cost plan. Computed at build from the same cost rules the calculator renders · ~ = proxy estimate

Provider / planPriceReq/$1
Ozore Pro$3512,048
Ozore Basic$1012,048
CommandCode Go$1~2,285
CommandCode GOAT$10~1,600
Z.ai GLM Coding Max$1681,029

“n/a” = the provider publishes no price structure for GLM-5.3-Flash — the calculator renders the same honest n/a. ◆ best value = highest computed Req/$1.

What real work costs

What GLM-5.3-Flash costs on real coding tasks

Illustrative cost of common coding tasks on the best-priced plans above, scaled from each plan's live per-request price to a task-typical token volume. Task token assumptions are shared ~proxy estimates — real tasks vary, and they are never per-provider shapes.

TaskOzore ProOzore BasicCommandCode GoCommandCode GOATZ.ai GLM Coding Max
Quick lookup / one-liner8,000 in · 1,000 out~1¢~1¢~1¢~1¢~1¢
Review a 500-line PR60,000 in · 4,000 out~1¢~1¢~1¢~1¢~1¢
Fix a bug (agent loop)180,000 in · 12,000 out~1¢~1¢~1¢~1¢~1¢
Refactor a module320,000 in · 20,000 out~1¢~1¢~3¢~3¢~1¢
Full-repo agent run900,000 in · 45,000 out~1¢~1¢~7¢~7¢~2¢

Per-task cost = the plan's effective 56,200-token request price scaled to each task's token volume at the same shape mix. ~ = proxy estimate — shared task-token assumptions, not a quote. “n/a” = GLM-5.3-Flash unpriced on that plan.

Common questions

How many requests per $1 do I get on GLM-5.3-Flash?

The best plan pricing GLM-5.3-Flash delivers 12,048 Req/$1 at the model's typical request shape — on Ozore Pro. Figures are exact confidence: they are computed from published rates and usage caps, proxy estimates are marked with a ~, and borrowed-cap plans use a published API cap at list price, so their Req/$1 is a value comparison rather than a measured allowance.

What's the cheapest plan with GLM-5.3-Flash?

CommandCode Go at $1/mo. The cheapest plan is not always the best value — check Req/$1 above.

What is GLM-5.3-Flash?

GLM-5.3-Flash is a confirmed Z.AI (Zhipu AI) GLM-family reasoning model designed for coding, sustained agentic work, and production workloads — the same model that launched on OpenRouter/CommandCode under the stealth codename "Ox Alpha." It is listed in Z.AI’s own "Latest Models" pricing (docs.z.ai/guides/overview/pricing) at $0.15 input / $0.075 cached / $0.50 output, with a 50% launch discount through 2026-09-09. Z.AI confirmed the ox-alpha identity to Bloomberg on 2026-08-26, and its weights are open (MIT). It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with visual context.

Who makes GLM-5.3-Flash?

GLM-5.3-Flash is made by Z.AI (Zhipu AI), a Beijing-based AI lab — the company behind the GLM family (GLM-5, GLM-5.2, GLM-5.3). It is a confirmed Z.AI model: Z.AI’s own "Latest Models" pricing (docs.z.ai/guides/overview/pricing) lists GLM-5.3-Flash at $0.15 input / $0.075 cached / $0.50 output, with a 50% launch discount through 2026-09-09. The "Ox Alpha" name used during the preview was a stealth codename; Z.AI confirmed the identity to Bloomberg on 2026-08-26 and the weights are open (MIT). GLM-5.3-Flash is a separate entry from the paid GLM-5.3 flagship ($1.4/$4.4). The OpenRouter/CommandCode route id remains stealth/ox-alpha.

Is GLM-5.3-Flash free?

Yes — during the preview window GLM-5.3-Flash is billed at $0.00/M for input, output, and cache reads. This is a temporary promotional preview, not a permanent free tier. The preview is expected to last roughly one week from launch (August 20, 2026).

What is the context window and output limit?

GLM-5.3-Flash has a 1,048,576-token context window (approximately 1M tokens) and supports up to 131,072 completion tokens per request.

Which CommandCode plans offer GLM-5.3-Flash?

GLM-5.3-Flash is available on CommandCode Go and above — Go ($1), GOAT ($10), Pro ($20), Provider ($15), Max 10× ($100), and Max 20× ($200). The $1 Go tier includes it despite being the entry-level plan.

Is GLM-5.3-Flash multimodal?

Yes — GLM-5.3-Flash accepts text, images, and video as input and returns text output. Note that Z.AI does not present GLM-5.3 as its native multimodal product (GLM-5V-Turbo is); this multimodal Flash variant is the reattributed "Ox Alpha." It supports function calling (tools/tool_choice) and structured outputs (response_format for JSON, without JSON-schema enforcement).

How do I access GLM-5.3-Flash?

On CommandCode, use the command `cmd --model stealth/ox-alpha` or type `/model` in a session and pick it. You can switch mid-session without losing context. OpenRouter also serves it as `stealth/ox-alpha` through an OpenAI-compatible API.

Does GLM-5.3-Flash beat GPT-5.6 Sol or Claude Fable 5?

Benchmarks are preliminary (10-task test, not audited) and LMArena has not yet scored GLM-5.3-Flash. Unverified community claims circulate, but tokenplans does not assert ranking superiority without verified leaderboard data. The model is positioned for long-horizon agentic work, not necessarily as a direct GPT-5.6 Sol or Claude Fable 5 competitor in all dimensions.

Newsletter

The weekly brief — plan churn, dated.

A plain weekly digest of every tracked price move, ceiling change and model-list update — pulled straight from the dated changelog ledger, with a one-line “worth switching?” note on each. Free. No spam, unsubscribe anytime.

Free · weekly · unsubscribe anytime