Did this plan quietly get worse?

Plan events, dated and sourced.

Every price move, ceiling change, model-list update, launch and discontinuation we log for the tracked plans — plus the corrections we make to our own data. History is un-copyable backwards: the changelog starts at the typed data layer (2026-08-02) and grows from here. Subscribe via RSS for new entries.

Events logged

162

Providers tracked

31

Stability index

Plan events by provider · last 90 days

The price, ceiling, model, policy, launch, incident and discontinuation events logged in the window — a signal for activity, not a verdict. Segments show the direction: worse,better, orneutral / announced. Read the timeline below for detail.

ProviderEvents / 90dReading
CommandCode23
2 worse16 better5 neutral
OpenCode19
6 worse13 better
Ozore14
7 better7 neutral
LLM Gateway DevPass13
9 better4 neutral
Nous Portal13
1 worse9 better3 neutral
Cursor10
7 better3 neutral
OpenAI10
5 better5 neutral
Cline5
1 worse3 better1 neutral
QwenCloud5
2 better3 neutral
Z.ai5
2 better3 neutral
X.AI4
2 better2 neutral
Chutes3
1 better2 neutral
Claude3
1 better2 neutral
GitHub Copilot3
2 better1 neutral
NanoGPT3
1 worse1 better1 neutral
Ollama3
2 better1 neutral
Phoenix Grove3
2 better1 neutral
Xiaomi MiMo3
2 better1 neutral
Yolo-Auto3
1 worse2 neutral
Google AI2
1 better1 neutral
Muse Code2
1 better1 neutral
Neural Watt2
1 better1 neutral
Standard Compute2
2 neutral
Synthetic2
1 better1 neutral
ZenMux2
2 neutral
BytePlus1
1 neutral
camelAI1
1 neutral
Kimi1
1 neutral
MiniMax1
1 neutral
StepFun1
1 neutral
Mistral0

Timeline

All logged events

  1. 25 Sept 2026

    • CeilingMuse Code

      Meta’s own subscription docs and product page publish the Muse Code usage multiples as 5× (High Usage) and 20× (Power Usage) of the Everyday Usage plan — the records stored 3× and 10× from the 2026-09-01 press report. The High 5h ceiling moves 150 → 250 prompts and Power 500 → 1000; Everyday stays 10–50 prompts/5h, now worded with the provider’s “including image and video uploads”, and the three display names become Everyday Usage / High Usage / Power Usage. Prices ($5 / $15 / $50 per month) and the per-token proxy rates are unchanged.

      High Usage 5h cap 150 prompts (3× Everyday) · Power Usage 5h cap 500 (10× Everyday) — the launch-coverage multipliers the plan records stored → High Usage 5h cap 250 prompts (5× Everyday) · Power Usage 5h cap 1000 (20× Everyday) · Everyday Usage 10–50 prompts/5h incl. image & video uploads

      Verified 2026-09-25 · source

    • PriceCommandCode

      Rate-provenance correction: the Sol entry now stores what this plan’s own page publishes. CommandCode’s pricing-limits row prices gpt-5.6-sol at $5.00 input / $30.00 output / $0.50 cache read / $6.25 cache write per 1M, so the retained OpenAI list is gone from the rule. Measured in the calculator: cost per request $0.0131 -> $0.03355 on GOAT, Pro, Max 10x, Max 20x and Provider, and Sol’s Req/$1 falls from 534 to 209 (GOAT), 305 to 119 (Pro) and 115 to 45 (Max 10x/20x). The row is byte-identical to this page’s GPT-5.5 row and is recorded exactly as published.

      GPT-5.6 Sol priced at the retained OpenAI list $2.00 / $10.00 per 1M (cached read $0.20, cache write $2.50) → GPT-5.6 Sol priced at the plan page’s own row $5.00 / $30.00 per 1M (cached read $0.50, cache write $6.25; over-272K tier $10/$45/$1/$12.50)

      Verified 2026-09-25 · source

    • PriceLLM Gateway DevPass

      Rate-provenance correction: the Sol entry now stores the plan’s own catalogue. The DevPass coding-models payload lists gpt-5.6-sol at inputPrice 4e-6 / outputPrice 2e-5 per token and the gateway’s catalogue publishes cachedInputPrice 4e-7 on the mapping that backs it, i.e. $4.00 / $20.00 / $0.40 per 1M, so Lite, Pro and Max stop borrowing the OpenAI list. Measured: cost per request $0.0131 -> $0.0262 and Req/$1 229 -> 114 (Lite, Pro) and 229 -> 115 (Max).

      GPT-5.6 Sol priced at the retained OpenAI list $2.00 / $10.00 per 1M (cached read $0.20) → GPT-5.6 Sol priced at the plan’s own gateway row $4.00 / $20.00 per 1M (cached read $0.40)

      Verified 2026-09-25 · source

    • PriceOpenAI

      Rate-provenance correction: OpenAI’s pricing page headlines gpt-5.6-sol in its Standard table at $4.00 input / $20.00 output per 1M with cached input $0.40 and cache write $5.00, and notes that this promotional pricing is available at least through November 21, 2026. The $2.00 / $10.00 figure these three consumer rules carried is the page’s Batch and Flex tariff, byte-identical to its gpt-6-sol row. Measured: cost per request $0.0131 -> $0.0262, halving Sol’s Req/$1 on Plus (76 -> 38), Pro 5x (76 -> 38) and Pro 20x (76 -> 38).

      GPT-5.6 Sol comparability rate on ChatGPT Plus, Pro 5x and Pro 20x stored at $2.00 / $10.00 per 1M (cached read $0.20) → GPT-5.6 Sol comparability rate $4.00 / $20.00 per 1M (cached read $0.40, cache write $5.00) — the vendor’s own Standard row

      Verified 2026-09-25 · source

    • PriceGitHub Copilot

      Rate-provenance correction: the vendor block that prices Copilot’s third-party models cites the model owner’s own rate card, whose Standard table now reads $4.00 / $20.00 per 1M with cached input $0.40, so Pro+ and Max bill that instead of the retained list. Measured: cost per request $0.0131 -> $0.0262 and Req/$1 137 -> 68 (Pro+) and 153 -> 76 (Max). Sol’s rate on the Cursor plans was already the vendor’s $4/$20/$0.40 and does not move.

      GPT-5.6 Sol vendor list row stored at $2.00 / $10.00 per 1M (cached read $0.20) → GPT-5.6 Sol vendor list row $4.00 / $20.00 per 1M (cached read $0.40)

      Verified 2026-09-25 · source

  2. 24 Sept 2026

    • ModelsOllama

      Ollama Cloud adds deepseek-v4.1-flash to the model roster and to both tracked plans (Pro and Max) — the only roster change between the 2026-09-02 and 2026-09-24 pricing dumps, priced at the new pricing-table row. The page now publishes off-peak rates (0.5x outside Mon–Fri 12:00–18:00 UTC, all day weekends), documented in the notes of the Ollama DeepSeek cost rules.

      deepseek-v4.1-flash absent from the Ollama Cloud roster — Pro and Max list 19 cloud models → DeepSeek V4.1 Flash added to Pro and Max (deepseek-v4.1-flash: $0.30 input / $1.20 output / $0.006 cached input per 1M)

      Verified 2026-09-24 · source

    • PriceOllama

      Ollama withdraws the cached-input rates for mistral-large-3, nemotron-3-nano and qwen3.5:397b. What a subscriber pays is unchanged — the withdrawn rate equalled the input rate on all three, and the calculator bills cached tokens at the input rate both before and after — so the stored values are dropped without a replacement and the affected Req/$1 figures now carry the ≥ floor marker.

      Published cached-input rates equal to the input rate on three models: mistral-large-3 $0.50, nemotron-3-nano $0.06, qwen3.5:397b $0.60 → Cached-input cells show "-" — no published cached-read rate on mistral-large-3, nemotron-3-nano or qwen3.5:397b (input/output unchanged)

      Verified 2026-09-24 · source

    • ModelsOpenCode· Go

      The Go Prices table carries a GPT 6 Luna row (≤ 272K tokens) that the plan list passed over — it is a separate model from GPT 5.6 Luna, with its own row, its own model id (gpt-6-luna) and its own request-rate and endpoint entries, and OpenCode’s public catalog serves both. Added to the plan model list and priced at the page’s own rates with the page’s published request shape (1,000 input / 50,000 cached / 220 output). The Space Bunny Free row on the same table stays uncatalogued: it is a limited-time free promo with no published rate (see PRICING_SOURCES.md § OpenCode Go).

      OpenCode Go: GPT 6 Luna served on the page but absent from the plan model list and the calculator → OpenCode Go: GPT 6 Luna at the published ≤ 272K-token rates ($0.10/$0.50 per 1M, cached read $0.01, cache write $0.125, $15/mo cap)

      Verified 2026-09-24 · source

    • PriceStandard Compute

      Cache-provenance re-verify of the Standard Compute rule: the provider resells at list rates and publishes none of its own, so every row’s cached read must trace to the page it borrows from. GLM-5.3-Flash’s $0.075 was the expired 2026-09-09 Z.AI launch-promo figure and the GLM-5.3-FlashX row’s cache column — corrected to the $0.03 Z.AI prints for GLM-5.3-Flash today, which llmgateway’s zai mapping and the CommandCode row both carry. The other 13 cached reads matched their named sources exactly; Gemini 3.1 Pro’s citation moved off ai.google.dev/pricing, which now redirects through to a Google sign-in, to the Google Cloud Gemini pricing page. Input, output and budget unchanged.

      Standard Compute per-model rows borrowed at provider list rates; GLM-5.3-Flash’s cached read stored at $0.075 → GLM-5.3-Flash cached read $0.03 — the figure Z.AI publishes for GLM-5.3-Flash — on all four plans (Starter / Economy / Standard / Pro)

      Verified 2026-09-24 · source

    • PriceZenMux

      Cache-provenance re-verify of the ZenMux Builder Plan’s per-model rows against the provider’s own model API: every cached read on the rule now traces to a cache-read figure ZenMux prints, and four rows were priced below it. ZenMux has no cache-discount structure — its cache reads are the vendor list rates — so the stored 0.9x (Kimi K3, MiniMax M3) and 0.7x (Qwen3.8-Max) rows were not a ZenMux discount and were corrected, input and output with them; GPT-5.6 Sol was storing its 272K-token tier instead of the base tier. Every other row matches the page exactly.

      ZenMux PAYG rows stored with unsourced discounts: Kimi K3 $2.70/$13.50/$0.27 and MiniMax M3 $0.27/$1.08/$0.054 (both 0.9x the provider’s published rates), Qwen3.8-Max $1.40/$4.20/$0.119 (0.7x), and GPT-5.6 Sol at its 272K-token tier $4/$20/$0.40 → ZenMux rows at the rates the provider publishes on its own model API: Kimi K3 $3.00/$15.00/$0.30, MiniMax M3 $0.30/$1.20/$0.06, Qwen3.8-Max $2/$6/$0.25 (cache write $2.50) and GPT-5.6 Sol at its base tier $2/$10/$0.20 (cache write $2.50)

      Verified 2026-09-24 · source

    • CeilingOpenCode· Go

      The Go Prices table prices GLM-5.3-Flash at a $60 monthly limit, four times the $15 the calculator stored — the row was never documented in the plan’s cap table. Request shape and rates unchanged.

      OpenCode Go: GLM-5.3-Flash monthly model limit $15 → OpenCode Go: GLM-5.3-Flash monthly model limit $60

      Verified 2026-09-24 · source

    • CeilingOpenCode· Go

      The Go Prices table now prices Qwen3.7 Max at a $30 monthly limit, half the $60 the calculator stored (and half the sibling Qwen3.8 Flash tier’s $30 sits level with it, no longer below). Rates and request shape unchanged.

      OpenCode Go: Qwen3.7 Max monthly model limit $60 → OpenCode Go: Qwen3.7 Max monthly model limit $30

      Verified 2026-09-24 · source

  3. 23 Sept 2026

    • PriceCommandCode

      Rate re-verify of the plan’s 74 priced models against the provider’s own row table: DeepSeek V4 Flash and its Vision variant were stored above the live off-peak rate (the row’s own priceChangeNote dates the cut to Aug 16), Gemini 3.7 Flash’s 50%-off deal is no longer on the page, and GLM-5.3-Flash’s cached read is the page’s $0.03 rather than the expired Z.AI promo figure. Every other row matches.

      DeepSeek V4 Flash and V4 Flash Vision billed off-peak $0.22/$0.66/$0.007; Gemini 3.7 Flash billed the 50%-off deal rate $0.75/$3.75/$0.075; GLM-5.3-Flash cached read $0.075 → DeepSeek V4 Flash and Vision off-peak $0.15/$0.60/$0.003 (peak $0.30/$1.20/$0.006); Gemini 3.7 Flash back at list $1.50/$7.50/$0.15; GLM-5.3-Flash cached read $0.03

      Verified 2026-09-23 · source

    • PriceLLM Gateway DevPass

      Calculator cache correction: the DevPass plan page publishes input/output only, but the gateway’s own catalogue publishes a cachedInputPrice per upstream mapping, so cached tokens bill at that published rate rather than at the input rate — and three rows that billed cache reads as free (a 100% discount) now carry the published rate.

      Cached read billed at the input rate on 6 rows; MiniMax M3, Nemotron 3 Ultra 550B and DeepSeek V4 Pro billed cached reads at 0 → Cached read billed at the gateway catalogue’s published cachedInputPrice on the mapping each row stores (0.20 / 0.20 / 0.01 / 0.50 / 0.0036 / 0.0028 restored; 0.06 / 0.10 / 0.0036 replace the zeros)

      Verified 2026-09-23 · source

    • PriceCommandCode

      Calculator coverage correction: the provider’s own row marks Laguna S 2.1 as a free deal (100% off, "free while capacity lasts"), so its rates are promotional rather than a published allowance and the 0/0/0 entry is dropped. The engine’s cost-per-request guard already rendered this row n/a, so no published figure moves; the table stops looking priced for a free promo model (Ozore free-row convention).

      Laguna S 2.1 carried a $0 / $0 / $0 rate entry → Laguna S 2.1 not priced — no rate entry; the row renders n/a

      Verified 2026-09-23 · source

    • ModelsSynthetic· Subscription

      The synthetic.new/pricing Standard-models table dropped the zai-org/GLM-5.2 row and gained deepseek-ai/DeepSeek-V4.1-Flash (Beta, 512k context) — every other row and every published rate is unchanged.

      Subscription served Kimi K3, GLM-5.2, Nemotron 3 Super 120B, Qwen3.8-27B, GPT-OSS 120B, GLM-4.7-Flash and GLM-5.3-Flash → GLM-5.2 dropped; DeepSeek V4.1 Flash added (Beta, 512k context) — roster now DeepSeek V4.1 Flash, Kimi K3, Nemotron 3 Super 120B, Qwen3.8-27B, GPT-OSS 120B, GLM-4.7-Flash and GLM-5.3-Flash

      Verified 2026-09-23 · source

    • ModelsCommandCode

      CommandCode lists stepfun/Step-5-Preview ("600B sparse-MoE agentic coding with 1M context") in its CLI models page, and the pricing-limits page publishes the row ($1.00 input / $2.70 output / $0.05 cache read per 1M, no cache-write band, single rate tier). The availability map sets every tier true including individual-go (minPlanName "Go"), so the model lands on all six tracked plans. Per-model allowance is the standing GOAT $20 / Pro $30 shape — the announcement's "2x usage in GOAT" is not published anywhere and is not encoded.

      Go, GOAT, Pro, Max 10x, Max 20x and Provider served the StepFun family as Step 3.7 Flash and Step 3.5 Flash only → Step 5 Preview added to Go, GOAT, Pro, Max 10x, Max 20x and Provider at $1.00/$2.70 per 1M (cache read $0.05; per-model allowance $20 GOAT / $30 Pro)

      Verified 2026-09-23 · source

    • ModelsCommandCode

      CommandCode lists xai/grok-4.7 in its CLI models page and the pricing-limits page publishes the row with a launch deal (id grok-4.7-40-off, 40% off, starts 2026-09-21, expires 2026-09-27): deal-applied inputCost 1.2 / outputCost 3.6 / cacheReadCost 0.3 per 1M against list $2.00/$6.00/$0.50, with the >200K-token band at $2.40/$7.20/$0.60. The availability map sets individual-go false, so the model is added from GOAT up (Go keeps Grok 4.5 only) and joins COMMANDCODE_GO_EXCLUDED. Category opensource, so Max 10x/20x bill it against the standard pool. The GOAT and Pro plan pages carry the boosted allowance in their Monthly credits column ("$20 $35 through Sep 27th" / "$30 $45 through Sep 27th"), stored as the live figure like the MiMo V2.6 Flash and DeepSeek V4.1 Flash promo rows.

      GOAT, Pro, Max 10x, Max 20x and Provider served Grok 4.5 and Grok 4.6 only → Grok 4.7 added to GOAT, Pro, Max 10x, Max 20x and Provider at the launch-deal rate $1.20/$3.60 per 1M (cache read $0.30, 40% off list $2.00/$6.00/$0.50); per-model allowance $35 GOAT / $45 Pro for the deal window, up from the standing $20 / $30

      Verified 2026-09-23 · source

    • ModelsCursor

      cursor.com/docs/models-and-pricing lists /docs/models/grok-4-7 and states that the Cursor Models pool "includes Grok 4.7, Grok 4.6, Grok 4.5, and Composer 2.5" with "significantly more included usage for Grok 4.7". Its model-pricing table gives the base Grok 4.7 row $2 input / $0.50 cache read / $6 output — the same row Grok 4.6 carries — while the 500K-context and Fast variants ($4/$1/$12 and $6/$1.5/$18) are separate rows, not modeled here. The Cursor Models pool publishes no dollar size, so the calculator keeps pricing it against the Other Models pool budget with that caveat, as it does for Grok 4.5/4.6.

      Cursor Pro, Pro+ and Ultra served Grok 4.5, Grok 4.6 and Composer 2.5 from the Cursor Models pool → Grok 4.7 added to Cursor Pro, Pro+ and Ultra at the published Cursor Models pool rate $2.00/$6.00 per 1M (cache read $0.50)

      Verified 2026-09-23 · source

    • ModelsOpenCode

      opencode.ai/docs/go now lists Grok 4.7 first in its model list and its usage-limits table prices it at the same rates and $15 monthly model limit as Grok 4.6, with the same published request shape ("Grok 4.7/4.6 — 390 input, 32,500 cached, 120 output tokens per request"); the model also appears in opencode.ai/zen/go/v1/models and its data-policy row carries 30-day retention. The page publishes two bands (>200K at $4/$12/$1.00); the cheap band is stored, matching the Grok 4.6 convention.

      OpenCode Go served Grok 4.6 only from the Grok line → Grok 4.7 added to Go at $2.00/$6.00 per 1M (cache read $0.50, $15/mo model cap, ≤ 200K-token band)

      Verified 2026-09-23 · source

    • ModelsLLM Gateway DevPass

      The DevPass coding-models catalogue lists id grok-4-7 (family xai, premium: false, contextSize 500000, discount 0) with inputPrice 2e-6 / outputPrice 6e-6 per token = $2/$6 per 1M. All three tracked plans expose the same catalogue and differ only in allowance, so the model is added to Lite, Pro and Max. The catalogue publishes no cached rate, so cached read uses the vendor cache-read rate $0.50.

      DevPass Lite, Pro and Max served Grok 4.6 only from the Grok line → grok-4-7 added to DevPass Lite, Pro and Max at the catalogue rate $2.00/$6.00 per 1M (cache read $0.50)

      Verified 2026-09-23 · source

    • ModelsOzore

      The public supported-models endpoint lists x-ai/grok-4.7 and the public credits-pricing table bills it at list_input 2.00 / list_output 6.00 with a 0.25 discount, i.e. billed_input 1.50 / billed_output 4.50 per 1M — the same 25% off the maker list the other xAI rows on this table carry. Ozore publishes no cached-read rate, so the calculator estimates it at 10% of the input rate, this provider convention.

      Ozore Basic and Pro served Grok 4.6, Grok 4.5, Grok 4.3 and Grok Build → Grok 4.7 added to Ozore Basic and Pro at the billed rate $1.50/$4.50 per 1M (25% off list $2.00/$6.00)

      Verified 2026-09-23 · source

    • ModelsNous Portal

      The Nous Portal inference catalogue lists x-ai/grok-4.7 with prompt $1.00 / completion $3.00 / input_cache_read $0.25 per 1M on a 500K context, and carries an `original` block at xAI list $2.00/$6.00/$0.50 — i.e. a 50% launch discount. This rule stores live top-level rates (deal-applied where a promo is running), so the discounted figure is stored with the list rate noted; re-verify the row after the promo ends.

      Nous Portal Plus, Super and Ultra served Grok 4.5 and Grok 4.6 → x-ai/grok-4.7 added to Nous Portal Plus, Super and Ultra at the live catalogue rate $1.00/$3.00 per 1M (cache read $0.25, 50% off the $2.00/$6.00/$0.50 list)

      Verified 2026-09-23 · source

    • ModelsCommandCode

      CommandCode added the MiMo V2.6 generation to its model table: the CLI models page lists xiaomi/mimo-v2.6-pro, xiaomi/mimo-v2.6-flash and xiaomi/mimo-v2.6-pro-ultraspeed, and the pricing-limits page publishes each with input/output/cache-read costs and a per-tier availability map. Pro and Flash are true on individual-go and above; UltraSpeed stays false on individual-go, so the calculator adds it from GOAT up. Per-model credit allowances: Flash $67 GOAT / $77 Pro (a 72-hour launch boost through 2026-09-24, standing $30 / $40), Pro $20 / $30, UltraSpeed $10 / $20.

      Go, GOAT, Pro, Max 10x, Max 20x and Provider served MiMo V2.5 and MiMo V2.5 Pro only → MiMo V2.6 Pro ($0.435/$0.87, cache read $0.0036) and MiMo V2.6 Flash ($0.14/$0.28, cache read $0.0028) added to every tier; MiMo V2.6 Pro UltraSpeed ($4.35/$8.70, cache read $0.036) added to GOAT and above (not Go)

      Verified 2026-09-23 · source

    • ModelsOpenCode

      The OpenCode Go model list now includes MiMo-V2.6-Flash and MiMo-V2.6-Pro, and the page own usage-limits table prices both per model with the same monthly dollar caps the V2.5 pair carries. Request shapes read from the page: Flash 830 input / 71,500 cached / 295 output, Pro 790 / 86,000 / 305. The calculator resolves a model request shape globally, so the Xiaomi Token Plan MiMo V2.6 Flash rows — which publish no shape of their own — now inherit this shape instead of the generic default, exactly as the Xiaomi MiMo V2.5 rows already inherit the OpenCode V2.5 shape.

      OpenCode Go served MiMo-V2.5 and MiMo-V2.5-Pro only → MiMo-V2.6-Flash added at $0.14/$0.28 per 1M (cache read $0.0028, $60 monthly model limit) and MiMo-V2.6-Pro at $0.435/$0.87 (cache read $0.003625, $15 monthly limit)

      Verified 2026-09-23 · source

    • ModelsLLM Gateway DevPass

      The DevPass coding-models catalogue now lists mimo-v2.6-flash (recommended) and mimo-v2.6-pro in its Recommended tab, so both are added to all three tracked DevPass plans, which expose the same catalogue and differ only in the usage allowance. The live rows publish inputPrice 1.4e-7 / outputPrice 2.8e-7 per token for Flash and 4.35e-7 / 8.7e-7 for Pro; the catalogue publishes no cached rate, so the calculator uses each model own vendor cache-read rate ($0.0028 and $0.0036). The $4/$20 Flash figure read when the sweep bead was filed was an upstream data-entry slip, now corrected on the provider page.

      DevPass Lite, Pro and Max served no MiMo model (MiMo V2.5 Pro had left the catalogue) → MiMo V2.6 Flash and MiMo V2.6 Pro added to Lite, Pro and Max at the catalogue rates $0.14/$0.28 and $0.435/$0.87 per 1M

      Verified 2026-09-23 · source

    • ModelsOzore

      Ozore routes the MiMo V2.6 pair through both tracked plans: the public supported-models endpoint lists xiaomi/mimo-v2.6-pro and xiaomi/mimo-v2.6-flash alongside the V2.5 rows, and the public credits-pricing table bills them at the same rates as their V2.5 equivalents ($0.3045/$0.609 for Pro, $0.098/$0.196 for Flash). Ozore publishes no cached-read rate, so the calculator estimates it at 10% of the input rate, the convention used for the rest of the Ozore rows.

      Ozore Basic and Pro served MiMo V2.5 Pro and MiMo V2.5 only → MiMo V2.6 Pro and MiMo V2.6 Flash added to Ozore Basic and Pro at the billed rates $0.3045/$0.609 and $0.098/$0.196 per 1M (30% off list)

      Verified 2026-09-23 · source

    • ModelsNous Portal

      The Nous Portal inference catalogue lists xiaomi/mimo-v2.6-pro, xiaomi/mimo-v2.6-flash and xiaomi/mimo-v2.6-pro-ultraspeed (canonical slugs dated 20260921) with published input_cache_read rates of $0.0036, $0.0028 and $0.036, so all three are added to the shared Plus rule that Super and Ultra derive from.

      Nous Portal Plus, Super and Ultra served MiMo V2.5 and MiMo V2.5 Pro only → MiMo V2.6 Pro, MiMo V2.6 Flash and MiMo V2.6 Pro UltraSpeed added to Nous Portal Plus, Super and Ultra at $0.435/$0.87, $0.14/$0.28 and $4.35/$8.70 per 1M

      Verified 2026-09-23 · source

    • ModelsOpenAI

      OpenAI brought the GPT-6 mid and low tiers to Codex the day after their API launch: the developers.openai.com Codex pricing table now lists GPT-6 Sol and GPT-6 Luna on Plus, Pro 5x and Pro 20x, while the Free and Go cards show GPT-6 Luna alone at Standard speed in the desktop app, subject to rollout, so Go is the only tracked plan that does not serve Sol. Message ranges read off the live table; token rates are the OpenAI API list rates ($2/$10 for Sol, $0.10/$0.50 for Luna) used as the comparability proxy for these consumer plans.

      ChatGPT Go served GPT-5.6 Luna only; Plus, Pro 5x and Pro 20x topped out at GPT-6 Astra → GPT-6 Sol and GPT-6 Luna added to ChatGPT Plus, Pro 5x and Pro 20x (Sol 15-150 / 70-700 / 300-3,000 messages per 5h; Luna 350-3,000 / 1,750-14,000 / 7,000-56,000); Go gets GPT-6 Luna only

      Verified 2026-09-23 · source

    • ModelsCommandCode

      CommandCode now serves both GPT-6 tiers. Its pricing-limits page publishes gpt-6-sol at $2/$10 with cache read $0.20 and cache write $2.50 in the premium category, and the model availability map marks individual-pro, individual-pro-v1, individual-provider, individual-max and individual-ultra true while individual-go and individual-goat stay false, so the calculator adds Sol to Pro and above only. gpt-6-luna is listed at $0.10/$0.50 (cache read $0.01, cache write $0.125) in the opensource category and is true on every plan including Go and GOAT. Requests above 272K tokens bill the long-context band ($4/$15 for Sol, $0.20/$0.75 for Luna), matching the page own tier table.

      Pro, Max 10x, Max 20x and Provider topped out at GPT-5.6 Sol; Go and GOAT had no GPT-6 model at all → GPT-6 Sol added to Pro, Max 10x, Max 20x and Provider at $2 input / $10 output per 1M (cache read $0.20, cache write $2.50, long-context band $4/$15); GPT-6 Luna added to every tier at $0.10/$0.50 (cache read $0.01, cache write $0.125); Go and GOAT serve Luna only

      Verified 2026-09-23 · source

    • ModelsLLM Gateway DevPass

      The DevPass coding-models catalogue now lists gpt-6-sol and gpt-6-luna (both premium: false, 1.05M context), so both are added to all three tracked DevPass plans, which expose the same catalogue and differ only in the usage allowance. The live catalogue publishes inputPrice 2e-6 / outputPrice 1e-5 per token for Sol and 1e-7 / 5e-7 for Luna; it publishes no cached rate, so the calculator uses each model own vendor cache-read rate ($0.20 and $0.01).

      GPT-6 Astra was the newest OpenAI model on DevPass Lite, Pro and Max → GPT-6 Sol and GPT-6 Luna added to DevPass Lite, Pro and Max at the catalogue rates $2/$10 and $0.10/$0.50 per 1M

      Verified 2026-09-23 · source

    • ModelsOzore

      Ozore routes both GPT-6 tiers through its tracked plans: the public supported-models endpoint lists openai/gpt-6-sol and openai/gpt-6-luna, and the public credits-pricing table bills them at $1.30/$6.50 and $0.065/$0.325 per 1M, 35% off the OpenAI list rates. Ozore publishes no cached-read rate, so the calculator estimates it at 10% of the input rate, the convention used for the rest of the Ozore rows.

      GPT-6 Astra Pro was the only GPT-6 model on Ozore Basic and Pro → GPT-6 Sol and GPT-6 Luna added to Ozore Basic and Pro at the billed rates $1.30 input / $6.50 output and $0.065 / $0.325 per 1M (35% off list)

      Verified 2026-09-23 · source

    • ModelsNous Portal

      The Nous Portal inference catalogue lists openai/gpt-6-sol and openai/gpt-6-luna (canonical slugs dated 20260922) at $2/$10 and $0.10/$0.50 per 1M with published input_cache_read rates of $0.20 and $0.01, so both models are added to the shared Plus rule that Super and Ultra derive from. The catalogue also serves reasoning-mode pro variants of both; those are not separate tracked models, since OpenAI own rate card does not price them separately.

      GPT-6 Astra was the newest OpenAI model on Nous Portal Plus, Super and Ultra → GPT-6 Sol and GPT-6 Luna added to Nous Portal Plus, Super and Ultra at $2 input / $0.20 cached input / $10 output and $0.10 / $0.01 / $0.50 per 1M

      Verified 2026-09-23 · source

    • ModelsCommandCode

      CommandCode now serves Claude Opus 5.5 on its three upper tiers. The provider pricing-limits page publishes a $4/$20 per 1M rate with cache read $0.20 and cache write $5, and its availability map marks the model on individual-provider, individual-max and individual-ultra while leaving individual-go, individual-goat and individual-pro false, so only Max 10x, Max 20x and Provider carry it in the calculator. The rate is Anthropic list, matching the 40% cost reduction against Claude Opus 5.

      Max 10x, Max 20x and Provider listed Claude Opus 4.8 and Claude Opus 5 as their newest Opus models → Claude Opus 5.5 added to Max 10x, Max 20x and Provider at $4 input / $20 output per 1M (cache read $0.20, cache write $5); Go, GOAT and Pro do not serve it

      Verified 2026-09-23 · source

    • ModelsCursor

      Cursor documented Claude Opus 5.5 as a first-class model, so it is added to all three tracked Cursor plans, which share one third-party model table and differ only in the pool budget. Cursor publishes the model at Anthropic list ($4/$20) with a $5 cache write and $0.20 cache read; the model draws from the shared Other Models pool like the other Claude rows on the plan.

      Claude Opus 5 was the newest Opus model on Cursor Pro, Pro+ and Ultra → Claude Opus 5.5 added to Cursor Pro, Pro+ and Ultra at $4 input / $5 cache write / $0.20 cache read / $20 output per 1M

      Verified 2026-09-23 · source

    • ModelsLLM Gateway DevPass

      The DevPass coding-models catalogue now lists claude-opus-5-5 (premium: true, 1M context), so the model is added to all three tracked DevPass plans, which expose the same catalogue and differ only in the usage allowance. A re-read of the live catalogue on 2026-09-23 shows inputPrice 4e-6 / outputPrice 2e-5 per token, i.e. $4/$20 per 1M - the $2/$10 figure recorded when this sweep was filed was stale. The catalogue publishes no cached rate, so the calculator uses the vendor cache-read rate of $0.20.

      Claude Opus 5 was the newest Opus model on DevPass Lite, Pro and Max → Claude Opus 5.5 added to DevPass Lite, Pro and Max at the catalogue rate $4/$20 per 1M

      Verified 2026-09-23 · source

    • ModelsOzore

      Ozore routes Claude Opus 5.5 through both tracked plans: the public supported-models endpoint lists anthropic/claude-opus-5.5 in its premium tier and the public credits-pricing table bills it at $2.60/$13.00 per 1M, 35% off Anthropic list. Ozore publishes no cached-read rate, so the calculator estimates it at 10% of the input rate, the convention used for the rest of the Ozore rows.

      Claude Opus 5 was the newest Opus model on Ozore Basic and Pro → Claude Opus 5.5 added to Ozore Basic and Pro at the billed rate $2.60 input / $13.00 output per 1M (35% off list $4/$20)

      Verified 2026-09-23 · source

    • ModelsNous Portal

      The Nous Portal inference catalogue lists anthropic/claude-opus-5.5 (canonical slug anthropic/claude-opus-5.5-20260921) at $4 input / $20 output per 1M with a published input_cache_read of $0.20, so the model is added to the shared Plus rule that Super and Ultra derive from. Unlike the older Claude Opus rows on this rule, the cached rate is the catalogue own input_cache_read rather than a proxy multiple of another provider rate.

      Claude Opus 5 was the newest Opus model on Nous Portal Plus, Super and Ultra → Claude Opus 5.5 added to Nous Portal Plus, Super and Ultra at $4 input / $0.20 cached input / $20 output per 1M

      Verified 2026-09-23 · source

    • ModelsClaude

      Anthropic own subscription tiers now carry Claude Opus 5.5: the claude.com/pricing Models grid marks Opus as included on Pro, Max 5x and Max 20x, and the same page API tab publishes the model at $4 input / $20 output per 1M with $0.20 cached read and a $5 cache write. The tracked plan records had not yet been re-checked against the model, so this entry adds it to all three, priced with the Anthropic list rates the calculator already uses as the comparability proxy for these consumer plans.

      Claude Pro listed Fable 5.1, Fable 5 and Sonnet 5; Max 5x and Max 20x listed Fable 5.1, Fable 5 and Opus 4.8 → Claude Opus 5.5 added to Claude Pro, Max 5x and Max 20x at the Anthropic API list rate $4 input / $20 output / $0.20 cached read per 1M

      Verified 2026-09-23 · source

    • ModelsXiaomi MiMo

      Xiaomi’s Token Plan Individual Edition documentation now states that every tier covers the V2.6 flagship models, so all four plan model lists moved from the V2.5 generation to the V2.6 series and the calculator prices both V2.6 models on each tier. The same page confirms that mimo-v2.5-pro and mimo-v2.5 are taken offline at 10:00 Beijing time on 2026-10-21, and that mimo-v2.6-pro-ultraspeed is not obtainable through the Token Plan at all — tier upgrades do not unlock it.

      All four Xiaomi Token Plan tiers listed only MiMo V2.5 Pro and MiMo V2.5 → Lite / Standard / Pro / Max now list MiMo V2.6 Pro and MiMo V2.6 Flash alongside the retiring V2.5 pair, with per-model deduction rows (V2.6 Pro 2.5/300/600, V2.6 Flash 2/100/200 Credits per cache-hit / cache-miss / output token)

      Verified 2026-09-23 · source

    • ModelsClaude

      Anthropic released Claude Opus 5.5 on 2026-09-22 and its own models overview and rate card now list it at $4/$20 per 1M tokens (claude-opus-5-5, adaptive thinking, default effort medium) — about 40% cheaper to run than Claude Opus 5 at $5/$25, at Claude Fable 5.1 performance level on most work. The model enters the registry with a model page; the tracked Claude subscription tiers have not yet been re-checked against it, so no plan model list, price or cap changed in this entry.

      Claude Opus 5 was the newest Opus-line model in the registry → Claude Opus 5.5 added to the registry and given a model page (API claude-opus-5-5); no tracked Claude plan prices it yet

      Verified 2026-09-23 · source

    • ModelsOpenAI

      OpenAI released the GPT-6 mid and low tiers on 2026-09-22: GPT-6 Sol (gpt-6-sol) at $2/$10 per 1M tokens with cached input $0.20 and cache writes $2.50, and GPT-6 Luna (gpt-6-luna) at $0.10/$0.50, both with a 1.05M-token context and 128K max output. Astra stays the $10/$50 frontier tier above them. Both identifiers are live on OpenAI’s own model and pricing pages; they enter the registry here, and the plans that serve them are being checked provider by provider, so no plan model list, price or cap changed in this entry.

      GPT-6 Astra was the only GPT-6-generation model in the registry → GPT-6 Sol and GPT-6 Luna added to the registry, with a model page for GPT-6 Sol; no tracked OpenAI plan prices either yet

      Verified 2026-09-23 · source

    • ModelsX.AI

      xAI released Grok 4.7 on 2026-09-21 and its own model docs now list it as the flagship for code and everything else: 500,000-token context, configurable reasoning, $2 per 1M input and $6 per 1M output tokens — the same base rates as Grok 4.6. The model enters the registry with a model page; xAI’s own SuperGrok tiers and the tracked BYOK plans that already list it are being verified separately, so no plan model list, price or cap changed in this entry.

      Grok 4.6 was the newest Grok model in the registry → Grok 4.7 added to the registry and given a model page (API grok-4.7); no tracked xAI plan prices it yet

      Verified 2026-09-23 · source

    • ModelsXiaomi MiMo

      Xiaomi released the MiMo V2.6 series on 2026-09-22: mimo-v2.6-pro (trillion-parameter omni-modal flagship reasoning model), mimo-v2.6-flash (full-modality low-cost reasoning tier) and mimo-v2.6-pro-ultraspeed (Pro performance served up to 20x faster). All three enter the registry, with a model page for the Pro tier. Xiaomi’s own Token Plan page already advertises V2.6 flagship access on every tier while the stored Xiaomi plan records still list the V2.5 generation — that plan-data refresh is tracked separately, so no plan model list, price or cap changed in this entry.

      MiMo V2.5 and MiMo V2.5 Pro were the newest Xiaomi models in the registry → MiMo V2.6 Pro, MiMo V2.6 Flash and MiMo V2.6 Pro UltraSpeed added to the registry, with a model page for MiMo V2.6 Pro; the Xiaomi Token Plan model lists still show the V2.5 generation

      Verified 2026-09-23 · source

    • PriceOzore

      Calculator cache correction: Ozore publishes no cached-read rate, so cached tokens now bill at the input rate instead of an invented 10%-of-input estimate. Req/$1 is a conservative floor — GPT-6 Astra drops from ~47 to ~6 req/$1.

      Cached read estimated at 10% of the input rate (invented — provider publishes none) → Cached read billed at the input rate (no published cache price)

      Verified 2026-09-23 · source

    • PriceNous Portal

      Calculator cache correction: Nous publishes no cache-read rate on the rows without an input_cache_read field, so cached tokens now bill at the input rate (no assumed 0.8× proxy). Rows with a published cache rate (GPT-6 Sol/Luna, Claude Opus 5.5, Grok 4.7, MiMo V2.6) are unchanged.

      Cached read proxied at 0.8× the OpenRouter cache-read rate on rows with no published input_cache_read → Cached read billed at the input rate on those rows; published input_cache_read rows unchanged

      Verified 2026-09-23 · source

    • PriceChutes

      Calculator cache correction: Chutes publishes no cached-read rate, so cached tokens now bill at the input rate instead of an invented 50%-of-input estimate.

      Cached read estimated at 50% of the input rate (invented — provider publishes none) → Cached read billed at the input rate (no published cache price)

      Verified 2026-09-23 · source

    • PriceNeural Watt

      Calculator cache correction: Neural Watt publishes no cached-read rate, so cached tokens now bill at the fresh input rate instead of an invented 0.3× estimate.

      Cached read estimated at 0.3× the fresh rate (invented — provider publishes none) → Cached read billed at the fresh (energy-derived) rate

      Verified 2026-09-23 · source

    • PriceLLM Gateway DevPass

      Calculator cache correction: the LLM Gateway DevPass catalogue publishes no cached-read rate, so cached tokens now bill at the input rate instead of borrowing the vendor rate.

      Cached read borrowed from the upstream vendor cache-read rate → Cached read billed at the input rate (catalogue publishes no cache price)

      Verified 2026-09-23 · source

    • PolicyMiniMax

      Calculator request-shape correction: MiniMax M3 now uses OpenCode's published per-request estimate instead of the shared fallback — the same shape the MiniMax budget derivation already used. Moves the M3 Req/$1 hero and the MiniMax gap table; no plan price, cap or model list changed.

      MiniMax M3 request shape 850 fresh / 49,000 cached / 160 output (shared fallback) → MiniMax M3 request shape 510 fresh / 56,000 cached / 190 output (OpenCode published estimate)

      Verified 2026-09-23 · source

  4. 22 Sept 2026

    • PriceOzore

      Ozore rate-precision correction, not a provider move: after the coverage extension every served model prices from the billed-rate table at ozore.com/api/credits/pricing, but the 13 rows that predated it still carried the pricing hero table’s rounded display values. Six of those rows differed from what the provider actually bills, and all six are now normalized to their own billed row, so every stored Ozore rate traces to the machine-readable table and none is a rounded marketing figure. The deltas are small (largest: GLM-5.3-Flash input, $0.10 -> $0.0975, 2.5%) and mixed in sign — the hero table rounds up on some rows and down on others — so this is a precision fix rather than a systematic shift in Ozore value. No plan price, credit grant, model list or published rate changed; cached-read stays the site’s 10%-of-input estimate because Ozore publishes none.

      Six models priced from the pricing hero table’s rounded display values — Gemini 3.8 Flash and Gemini 3.7 Flash $0.98/$4.88, DeepSeek V4 Pro $0.28/$0.57, DeepSeek V4.1 Flash $0.11/$0.42, MiMo V2.5 Pro $0.30/$0.61, GLM-5.3-Flash $0.10/$0.33 per 1M → The same six at the billed rates their own rows publish — Gemini 3.8 Flash and Gemini 3.7 Flash $0.975/$4.875, DeepSeek V4 Pro $0.2828/$0.5655, DeepSeek V4.1 Flash $0.105/$0.42, MiMo V2.5 Pro $0.3045/$0.609, GLM-5.3-Flash $0.0975/$0.325 per 1M

      Verified 2026-09-22 · source

    • PriceOzore

      Ozore calculator coverage — the public rate table at ozore.com/api/credits/pricing publishes a billed input/output rate for every model the provider bills, so the per-model rules were extended from the 13 pricing-hero rows to the full table (36 models added, each rate traced to its own billed row). No plan price, credit grant or published rate changed. The six models served on the free tier only (Hy3, Nemotron 3 Ultra, Agnes 2.0 Flash, North Mini Code, Ling 3.0 Flash, GPT-OSS 120B) stay unpriced on purpose, because a $0 rate reads as an unmetered win.

      13 of 55 models per plan priced — the pricing-hero rows only, so 36 served models rendered n/a in the calculator → 49 of 55 models per plan priced from the public billed-rate table; only the 6 free-tier-only models stay unpriced

      Verified 2026-09-22 · source

    • ModelsOzore

      Ozore catalogue reconciliation — GPT-5.4 Mini was still stored from an earlier catalogue read but appears on no live Ozore surface: it is absent from the routing catalogue behind ozore.com/models (55 models), from the billed-rate table at ozore.com/api/credits/pricing (50 rows) and from its free tier (7 rows, rendered as the Free table on ozore.com/pricing-direct). The only surface still listing it, ozore.com/all-models, is a hand-written snapshot inside the site bundle that also describes the discontinued premium-token-pool and Max plans and is linked from no live page, so it is not treated as a current catalogue. GPT-OSS 120B, which that legacy page also lists, is deliberately kept: the same live credits-pricing free tier serves it as gpt-oss-120b (free-only, so it stays unpriced). No price, credit grant or published rate changed.

      56 models per plan — including GPT-5.4 Mini, which no live Ozore catalogue serves → 55 models per plan — GPT-5.4 Mini removed; the live routing catalogue (ozore.com/models, fed by ozore.com/api/supported-models) no longer lists it, and GPT-OSS 120B is kept because the live credits-pricing free tier still serves it

      Verified 2026-09-22 · source

  5. 21 Sept 2026

    • PriceQwenCloud

      Site data correction, not a claimed provider price cut: qwencloud.com/models/qwen3.8-max now publishes the model’s own rate row, so the calculator’s Qwen3.8-Max figure stops using the qwen3.7-max-preview proxy (which was cited as “3.8 not yet on the public PAYG table”) and roughly doubles Qwen3.8-Max Req/$1 on all four QwenCloud Token plans (Lite 37.5 → 71.75; the same correction applies to Essential, Standard and Pro). The provider’s internal per-token credit coefficients are console-only, so whether a subscriber’s credit consumption actually changed is not observable from the public pages — hence direction “neutral”.

      Qwen3.8-Max proxied from qwen3.7-max-preview: $2.5 input / $7.5 output / $0.5 implicit cache per 1M → Qwen3.8-Max at the provider’s own published row: $2 input / $6 output / $0.25 implicit cache per 1M (explicit-cache rows not modeled) — affects Lite, Essential, Standard and Pro, which inherit the per-model map

      Verified 2026-09-21 · source

    • LaunchQwenCloud· Essential Plan

      QwenCloud adds a new Essential tier to the Personal Token Plan at $16/mo with a limited-time $10/mo promo, slotting between Lite ($8, 2,500 credits/7d) and Standard ($25, 10,000 credits/7d) with 2.25× the Lite credit budget (5,625 credits/7d); the existing three tiers are unchanged.

      Three tiers: Lite ($8, 2,500 credits/7d), Standard ($25) and Pro ($80) → Essential Plan launched between Lite and Standard: $16/mo (limited-time $10/mo), 5,625 credits / 7 days, 2.25× the Lite credit budget, same model roster and Qwen Code access; Lite/Standard/Pro unchanged

      Verified 2026-09-21 · source

    • ModelsQwenCloud

      The docs "Supported models" allowlist for the Personal Token Plan grew from 7 to 10 text-capable models: Qwen3.8 Flash, GLM-5.3 and DeepSeek V4.1 Flash join the roster on Lite, Essential, Standard and Pro (the bare `auto` entry is a platform-side router, not a named model, and is deliberately not tracked; image/audio/video rows stay out of scope).

      7 models — Qwen3.8-Max, Qwen3.7-Max, Qwen3.7-Plus, Qwen3.6-Flash, GLM-5.2, DeepSeek V4 Pro, DeepSeek V4 Flash → 10 models — added Qwen3.8 Flash, GLM-5.3, DeepSeek V4.1 Flash

      Verified 2026-09-21 · source

    • PolicyQwenCloud

      The night-discount surfacing changed: the live docs list four night-discount models (up from the two stored) and the night rule itself is no longer labelled "limited time" — that wording now sits only on the promo prices in the pricing table. Discount terms (50% off Credits, 22:00–08:00 UTC+8) are unchanged.

      Night 22:00–08:00 UTC+8: 50% off Credits on qwen3.8-max and deepseek-v4-pro-0813 (limited time) → Night 22:00–08:00 UTC+8: 50% off Credits on qwen3.8-max, deepseek-v4-pro-0813, deepseek-v4-flash-0731 and deepseek-v4.1-flash

      Verified 2026-09-21 · source

    • ModelsOzore

      Ozore catalog addition — the supported-models catalog behind ozore.com/models now lists Ling 3.0 Flash as a standard (unlimited) model, so both Basic and Pro gain it. Prices ($10/$35), monthly credit grants ($20/$70) and every published per-token rate are unchanged.

      55 models per plan → 56 models per plan — added Ling 3.0 Flash (standard, unlimited on every plan)

      Verified 2026-09-21 · source

  6. 20 Sept 2026

    • ModelsCursor

      Cursor adds Muse Spark 1.3 to the model roster on Pro, Pro+ and Ultra at $1.25 input / $4.25 output / $0.15 cache read per million tokens (no cache write published); prices, pools and ceilings unchanged.

      Muse Spark 1.3 not listed on the Cursor model-pricing table → Muse Spark 1.3 added to the shared model roster on Pro, Pro+ and Ultra ($1.25 input / $4.25 output / $0.15 cache read)

      Verified 2026-09-20 · source

  7. 19 Sept 2026

    • ModelsNanoGPT· Subscription

      NanoGPT PRO added GLM 5.3 ("GLM 5 / 5.1 / 5.2 / 5.3" flagship bullet) and a new separate "GLM 5.3 Flash — Fast reasoning, vision, and agent model" bullet to the included text roster; both already exist in the site’s models registry. Price ($12/mo), the 60M included-input weekly cap and the 100 free images/day are unchanged.

      8 models — GLM-5.2, MiniMax M3, DeepSeek V4 Pro, GLM-5, GLM-5.1, MiMo V2.5, MiMo V2.5 Pro, Kimi K2.7 Code → 10 models — added GLM-5.3 (2x included-input multiplier) and GLM-5.3-Flash (1x)

      Verified 2026-09-19 · source

    • CeilingNanoGPT· Subscription

      The live page enumerates the 2x included-input family and adds that "This includes each family’s regular and Thinking variants"; GLM 5.3 joins the 2x set and GLM 5.3 Flash is 1x, so the stored rateLimit string and the calculator multipliers now match the page.

      2x included-input multiplier: GLM 5 / GLM 5.1 / Kimi K2.7 Code / DeepSeek V4 Pro → 2x included-input multiplier: GLM 5 / GLM 5.1 / GLM 5.3 (including GLM Latest) / Kimi K2.7 Code / DeepSeek V4 Pro / V4 Pro 0813 — 1x: GLM 5.2, GLM 5.3 Flash, MiMo V2.5 / V2.5 Pro, Kimi K2.5 / K2.6, MiniMax M3

      Verified 2026-09-19 · source

    • PolicyNanoGPT· Subscription

      Two policy items from the same page read, logged as one event. (1) The 5% PAYG discount scope narrowed from every non-included text model to an "eligible" subset with provider-side exclusions, so a subscriber can now pay full PAYG price on models the old copy covered. (2) The FAQ states web search is not included in subscription coverage, so the plan pricingNotes now surfaces it; that part is a surfaced caveat, not a change in what PRO delivers.

      5% PAYG discount on "all other text models" (FAQ: "All other text models that are not already included in the subscription. This includes TEE variants and proprietary models like ChatGPT (OpenAI), Gemini (Google), and Claude (Anthropic)."); pricingNotes said nothing about web search → 5% PAYG discount on "eligible paid text models" (FAQ: "Eligible paid text models that are not already covered by the subscription. This includes most TEE variants and proprietary models from OpenAI, Google, Anthropic, and others; exclusions may apply."); included-models card: "All included open-source text models" → "All included text models"; pricingNotes now records that web search is not covered by PRO

      Verified 2026-09-19 · source

  8. 18 Sept 2026

    • PriceChutes

      Chutes rate-table re-read, logged as one event for three rows. Kimi K2.6 (moonshotai/Kimi-K2.6-TEE) cut from $0.58/$3.40 to $0.50/$2.85 per 1M, better for Plus and Pro. DeepSeek V4 Flash (deepseek-ai/DeepSeek-V4-Flash-0731-TEE) went up from $0.14/$0.28 to $0.44/$1.32, a 3.1x/4.7x rise the stored record missed: the provider row already read $0.44/$1.32 in the 2026-08-28 dump, so it is both a provider rise and a stale-record correction. GLM-5.2 corrected from $1.40/$4.40 to $1.25/$3.95, a site-data fix rather than a provider move (the old figure was the z.ai list rate, not this chute row). All three move tokens-per-allowance-dollar on Plus and Pro, where usage is metered in PAYG-equivalent dollars at 5x the plan price; plan prices ($10/$20), quota and the 14-model rosters are unchanged.

      Kimi K2.6 $0.58/$3.40 · DeepSeek V4 Flash (V4-Flash-0731 TEE) $0.14/$0.28 · GLM-5.2 $1.40/$4.40 per 1M (Plus and Pro) → Kimi K2.6 $0.50/$2.85 · DeepSeek V4 Flash (V4-Flash-0731 TEE) $0.44/$1.32 · GLM-5.2 $1.25/$3.95 per 1M (Plus and Pro)

      Verified 2026-09-18 · source

    • ModelsCline· Pass

      ClinePass added DeepSeek V4.1 Flash and Muse Spark 1.3 Contributor to its plan roster, announced in the Cline team email of 2026-09-17. Neither model is listed on cline.bot/cline-pass or in the docs Reference-pricing table yet (checked 2026-09-18), so both are stored without a rate and their calculator rows show n/a. Price ($9.99/month), quota and all published reference rates are unchanged.

      14 models → 16 models — added DeepSeek V4.1 Flash and Muse Spark 1.3 Contributor

      Verified 2026-09-18 · source

  9. 17 Sept 2026

    • PriceCline· Pass

      ClinePass ended its limited-time first-month promotion: every price block on cline.bot/cline-pass (hero, pricing card, FAQ, CTA) now reads $9.99/month and the $4.99 intro offer is gone. The standard rate was already $9.99 and is unchanged — what disappeared is the $4.99 first-month price.

      $4.99 first month, then $9.99/mo (limited-time promotion) → $9.99/month from the first month (no promotion)

      Verified 2026-09-17 · source

    • ModelsCline· Pass

      The ClinePass roster grew from 12 to 14 models with no removals: the plan page now lists GLM 5.3 and GLM 5.3 Flash alongside the previous 12. Price, quota and all published reference rates are unchanged.

      12 models → 14 models — added GLM 5.3 and GLM 5.3 Flash

      Verified 2026-09-17 · source

    • ModelsYolo-Auto

      Site correction: the Builder and Pro records stored “Qwen 3.8 27B”, a checkpoint Yolo-Auto’s live /models table does not list. Both plans now carry the public model ID the provider actually serves, Qwen3.8 Flash (qwen3.8-flash, 256K context, free and paid); the paid yolo route is a configurable server-side target, so it is deliberately not stored as a model name. Plan prices ($19/$39), contexts and the flat-rate terms are unchanged.

      Qwen 3.8 27B → Qwen3.8 Flash (public model ID qwen3.8-flash)

      Verified 2026-09-17 · source

    • ModelsOzore

      Ozore catalog refresh — both Basic and Pro now list the live model set: Claude Fable 5.1, GPT-6 Astra, Gemini 3.8 Flash, DeepSeek V4.1 Flash and Muse Spark 1.3 Contributor were added, while Gemini 3 Flash, MiniMax M2.5, GPT-5 Mini and GPT-5 Nano are no longer served and were removed. Prices, monthly credit grants and existing rates are unchanged.

      54 models per plan (2026-08-31 catalog snapshot) → 55 models per plan — added Claude Fable 5.1, GPT-6 Astra, Gemini 3.8 Flash, DeepSeek V4.1 Flash and Muse Spark 1.3 Contributor; removed Gemini 3 Flash, MiniMax M2.5, GPT-5 Mini and GPT-5 Nano

      Verified 2026-09-17 · source

    • ModelsCommandCode

      Site correction: CommandCode plans advertised models a subscriber cannot reach. Go, GOAT and Pro priced premium models that need a higher plan — the Pro page states “Claude Opus, Claude Fable, and Fugu Ultra require a Max plan” — and Claude Sonnet 4.5 was still priced on Max 10×, Max 20× and Provider after CommandCode deprecated it. Those rows now render n/a instead of a request estimate. Rates, plan prices, credit grants, per-model allowances and window fractions are unchanged.

      Model rosters priced models CommandCode’s published availability map marks unavailable on that plan — Go and GOAT listed Fugu Ultra and Muse Spark 1.1, Go also priced GPT-5.6 Sol, Pro listed Claude Fable 5, Opus 5, Opus 4.7 and Opus 4.6, and every plan from Go through Max 10×/Max 20× and Provider still listed the deprecated Claude Sonnet 4.5 → Each plan prices only what the live availability map (individual-go / -goat / -pro / -max / -provider flags, cross-checked against each model’s minPlanName on the plan pages) grants it: Go drops Fugu Ultra, Muse Spark 1.1 and GPT-5.6 Sol; GOAT drops Fugu Ultra and Muse Spark 1.1; Pro drops Claude Fable 5, Opus 5, Opus 4.7, Opus 4.6 and Fugu Ultra; Claude Sonnet 4.5 is gone from every plan (deprecated provider-side)

      Verified 2026-09-17 · source

    • CeilingCommandCode

      Site correction: CommandCode GOAT and Pro credits are per-model allowances, not one shared pool, so the calculator was overstating most rows — a model with a $20 allowance was modeled as if it could spend GOAT’s whole $70 (Pro’s $80). Every GOAT and Pro row now budgets min(plan pool, published allowance), which reproduces the vendor’s own request estimates (DeepSeek V4.1 Flash: 154,000 req/mo on GOAT at its $60 allowance). DeepSeek V4.1 Flash’s allowance is $40 GOAT / $50 Pro, boosted to $60/$70 through 2026-09-20. Rates, plan prices and credit grants are unchanged.

      Calculator rules modeled every GOAT model at the plan’s full $70 credit pool and every Pro model at $80 — CommandCode’s published per-model credit allowances were never applied → Per-model monthly credit allowances applied (a model may draw at most its own allowance from the pool): GOAT $20 standard, $30–$33 on the MiMo / Qwen3.6-Plus / Qwen3.7 rows, $40 on the Flash rows (Gemini 3.7/3.8 Flash, GLM-5.3 Flash), $47 MiniMax M3, $60 on the DeepSeek/Kimi-K2.7-Code rows, up to the full $70 on the negotiated rows (GPT-5.6 Sol, GLM-5.2, Tencent Hy3, Qwen 3.8 27B); Pro the same shape at $30 / $40–$43 / $50 / $57 / $70 / $80

      Verified 2026-09-17 · source

  10. 16 Sept 2026

    • CeilingPhoenix Grove

      Phoenix Grove raises the monthly token allowance on all six tiers (Taster 15M → 18M Flash tokens; Basic–Canopy standard + light windows up, e.g. Basic 50M/400M → 85M/450M, Canopy 800M/6B → 1.3B/6.6B). Prices and agent concurrency (3/5/6/7/8) unchanged.

      Monthly token allowances — Taster 15M Flash; Basic 50M + 400M light; Pro 100M + 800M light; Elite 200M + 1.6B light; Ultra 400M + 3B light; Canopy 800M + 6B light → Monthly token allowances — Taster 18M Flash; Basic 85M + 450M light; Pro 165M + 890M light; Elite 335M + 1.8B light; Ultra 665M + 3.3B light; Canopy 1.3B + 6.6B light

      Verified 2026-09-16 · source

    • ModelsPhoenix Grove

      Phoenix Grove refreshes its model catalog from the live bundle: GLM-5.3, GLM-5.3-Flash and DeepSeek V4.1 Flash join every Basic+ plan, MiniMax M2.7 / GLM-5 / GLM-4.7-Flash are dropped from the rows after disappearing from the live catalog, and Taster newly serves GLM 5.3 Flash alongside DeepSeek V4 Flash.

      Catalog: 17 canonical names incl. MiniMax M2.7, GLM-5 and GLM-4.7-Flash; Taster = DeepSeek V4 Flash only → Catalog: GLM-5.3, GLM-5.3-Flash and DeepSeek V4.1 Flash added to Basic+; MiniMax M2.7 / GLM-5 / GLM-4.7-Flash dropped (absent from the live catalog); Taster now serves DeepSeek V4 Flash + GLM 5.3 Flash

      Verified 2026-09-16 · source

    • PriceOzore

      Site correction: Ozore per-model value figures were overstated. Billing cached tokens at $0 assumed free cache — against the site cache-heavy agentic shape (~86K cached of ~87K input tokens) that inflated requests-per-dollar 3.7–7x on every Ozore row and put the plans at the top of the value tables on an assumption the provider never published. Cached-read is now estimated at 10% of the input rate (the convention already used for maker-rate rows, and the cache-hit ratio DeepSeek/OpenAI publish); because that is derived, both plans are proxy confidence and their numbers carry the estimate marker. The Grok 4.6 row was also refreshed from 35% off to the live 25% off. Plan price, monthly credit grant and the published input/output rates are unchanged.

      Ozore Basic / Pro value figures computed with cached-read billed at $0 (the provider publishes no cached-read rate) and the Grok 4.6 row at the stale 2026-08-31 rate ($1.30/$3.90 = 35% off) → Cached-read estimated at 10% of the input rate, Grok 4.6 refreshed to its live rate ($1.50/$4.50 = 25% off), and both plan rules moved to proxy confidence so every Ozore figure renders as an estimate

      Verified 2026-09-16 · source

  11. 15 Sept 2026

    • PolicyX.AI

      Grok Bot entitlement surfaced on the xAI rows (2026-09-15 sweep). The $30 SuperGrok card now advertises "Exclusive access to Grok Bot", and the Cursor Grok Bot matrix (individual SuperGrok, SuperGrok Plus, SuperGrok Heavy and X Premium+ qualify; SuperGrok Lite does not) confirms the benefit is a linked weekly Grok Bot usage grant, not a Cursor plan. SuperGrok Heavy no longer claims a free Cursor Ultra — cursor.com/help/grok-bot/supergrok-heavy now redirects to cursor.com/help/grok-bot/supergrok, which describes the link as a usage grant that does not stack with a Cursor plan and never lists Cursor Ultra as included. No price, cap, model list or calculator rate changed (plan-card prices re-verified $10/$30/$100/$300).

      Grok Bot entitlement not recorded on the xAI plan rows; SuperGrok Heavy carried an unverifiable free Cursor Ultra claim citing cursor.com/help/grok-bot/supergrok-heavy → Grok Bot recorded on SuperGrok ($30) and SuperGrok Heavy ($300): included from SuperGrok up and not on Lite, granted as weekly Grok Bot usage when the Grok or X account is linked; the SuperGrok Heavy Cursor Ultra claim removed

      Verified 2026-09-15 · source

    • LaunchcamelAI· Stream

      camelStream added — $5/mo per stream for unlimited tokens on an auto-routed fleet (GPT-5.6 Luna, Muse Spark 1.3, GLM-5.3 Flash, DeepSeek V4.1 Flash), OpenAI- and Anthropic-compatible API, one concurrent generation per stream, no token metering; prompts and outputs are licensed for AI training and ad targeting.

      Verified 2026-09-15 · source

    • PolicyZ.ai

      docs.z.ai/devpack/overview (Credit Calculation) publishes the time-of-day credit rule — "During off-peak hours, model usage is charged at 50% of the standard credit rate" with peak hours Monday to Friday, 14:00–18:00 SGT (UTC+8) — and it is now stored on the three plan rows as an off-peak tooltip. The 50% covers model usage only (MCP calls keep their own multipliers); no price, credit allowance or calculator rate changed.

      GLM Coding off-peak terms not published on the plan rows (the 2026-08-08 limits entry noted "off-peak usage costs 50%" with no peak window) → Off-peak rule surfaced on GLM Coding Lite / Pro / Max: model usage billed at 50% of the standard credit rate outside peak hours; peak = Monday–Friday 14:00–18:00 Singapore time (UTC+8)

      Verified 2026-09-15 · source

    • ModelsCommandCode

      The pricing-limits availability map lists meta/muse-spark-1.3 (individual-go false, GOAT and above true) at $1.25/$4.25/$0.15 and meta/muse-spark-1.3-contributor (every plan including Go) at $0.10/$0.20/$0.002; both rows added to the stored plan model arrays plus per-model calculator rates (1.3 excluded from the $1 Go tier).

      Muse Spark 1.1 · 1.2 · 1.2 Contributor on the CommandCode plan model lists (no Muse Spark 1.3 row) → Muse Spark 1.3 added to GOAT / Pro / Max 10× / Max 20× / Provider (GOAT and above) and Muse Spark 1.3 Contributor added to every plan, Go included

      Verified 2026-09-15 · source

    • ModelsOpenCode

      opencode.ai/docs/go publishes Muse Spark 1.3 Contributor at $0.10/$0.20/$0.002 (cached read $0.002) with its own request shape (620 in / 71,400 cached / 300 out) and rate limits of 45,300 / 113,300 / 226,600 requests per 5h / week / month; added to the Go plan model array and the USD rule map at the $60/mo per-model cap.

      Muse Spark 1.2 Contributor on the OpenCode Go model list (no Muse Spark 1.3 row) → Muse Spark 1.3 Contributor added to the OpenCode Go model list

      Verified 2026-09-15 · source

    • ModelsNous Portal

      Two models the Portal subscription already proxies were missing from the Nous model lists: DeepSeek V4.1 Flash (launched 2026-09-10; portal row at the DeepSeek off-peak headline $0.15/$0.60, cached read $0.015) and GLM-5.3-Flash ($0.075/$0.25, cached $0.015 at its 50% launch rate). Both added to Plus, Super and Ultra with per-model rate entries in the Nous credits rule.

      Nous Plus / Super / Ultra model lists: 27 models, no DeepSeek V4.1 Flash or GLM-5.3-Flash row → Nous Plus / Super / Ultra model lists: 29 models — DeepSeek V4.1 Flash ($0.15/$0.60/$0.015 per 1M) and GLM-5.3-Flash ($0.075/$0.25/$0.015 per 1M) added

      Verified 2026-09-15 · source

    • PriceNous Portal

      Site correction: Hy3 was stored from the 0.8×-OpenRouter proxy read of portal.nousresearch.com/info. Re-read from the portal's own inference catalog row (pricing.original 0.132/0.528/0.033 is the list rate) — input and cached marginally cheaper, output higher. Applies to Plus, Super and Ultra.

      Hy3: $0.11 in / $0.42 out / $0.0264 cached read per 1M → Hy3: $0.105 in / $0.435 out / $0.0263 cached read per 1M (portal inference-API top-level row)

      Verified 2026-09-15 · source

  12. 13 Sept 2026

    • CeilingOpenAI

      Live Codex usage-limits table re-read, logged as one event: Plus message ranges for Sol, Terra and Luna all grew (Luna 50–280 → 250–2,000) and the dedicated Pro 5× and Pro 20× columns are stored for the first time. Figures are the table's unlabelled message windows; the page's only window statement is the footnote "Weekly limits may also apply", so no per-5h unit is claimed.

      Codex caps 2026-08-24: Plus Sol 15–90 · Terra 20–110 · Luna 50–280 (Astra 5-45); Pro 5× / Pro 20× columns not stored (Plus-era figures shown instead) → Plus Sol 10–100 · Terra 25–200 · Luna 250–2,000 · Astra 5-45 (Luna now the highest, ~10× the next-highest top of range); Pro 5× Sol 50–500 · Terra 125–1,000 · Luna 1,250–10,000 · Astra 25–225; Pro 20× Sol 200–2,000 · Terra 500–4,000 · Luna 5,000–40,000 · Astra 100–900

      Verified 2026-09-13 · source

    • ModelsOpenAI

      Compare-table sync: GPT-5.5 is not a named row and sits under "Legacy models" = Go No, so it is dropped from Go (its Luna row is Yes) while staying on Plus (Legacy = Plus Yes, own 15-80 Codex column). Terra (Plus Yes / Pro Unlimited*) and Luna (Yes everywhere) added to Plus and both Pro tiers; GPT-5 Thinking Mini (Go Yes) not stored — not in the model registry.

      Go ['GPT-5.5'] · Plus ['GPT-6 Astra','GPT-5.6 Sol','GPT-5.5'] · Pro 5×/20× ['GPT-6 Astra','GPT-5.6 Sol Pro','GPT-5.6 Sol'] → Go ['GPT-5.6 Luna'] (GPT-5.5 dropped) · Plus + GPT-5.6 Terra + GPT-5.6 Luna · Pro 5×/20× + GPT-5.6 Terra + GPT-5.6 Luna

      Verified 2026-09-13 · source

  13. 12 Sept 2026

    • PriceOpenCode· Go

      Provider-side rate cut on the two older DeepSeek Flash rows — off-peak $0.22/$0.66/$0.007 → $0.15/$0.60/$0.003, the same off-peak rates the new DeepSeek V4.1 Flash launched at. Caps unchanged ($30/mo V4 Flash, $15/mo V4 Flash Vision Exp) and the docs request-estimate table now reads 13,000 / 32,500 / 65,000 and 6,500 / 16,250 / 32,500 req per 5h/week/month. CommandCode still lists V4 Flash Vision Exp at the old rates, so the two gateways now legitimately differ.

      OpenCode Go: DeepSeek V4 Flash and V4 Flash Vision Exp off-peak $0.22 in / $0.66 out / $0.007 cached read per 1M → OpenCode Go: both rows off-peak $0.15 in / $0.60 out / $0.003 cached read per 1M (peak 2× = $0.30/$1.20/$0.006)

      Verified 2026-09-12 · source

    • ModelsLLM Gateway DevPass

      DeepSeek V4.1 Flash (launched 2026-09-10) was already in the DevPass coding catalogue but missing from the Lite / Pro / Max headline model arrays, so the 2026-09-10 refresh left it out. Added to all three arrays with its own rate entry in the DevPass rule; headline arrays 15 → 16, modelsMore +108 → +109.

      DevPass Lite / Pro / Max headline arrays: no DeepSeek V4.1 Flash row → DevPass Lite / Pro / Max list DeepSeek V4.1 Flash at the gateway list rate $0.15 in / $0.60 out / $0.003 cached read per 1M

      Verified 2026-09-12 · source

    • ModelsOpenCode· Go

      DeepSeek V4.1 Flash (launched 2026-09-10) joined the OpenCode Go lineup at a $15/mo usage cap — half the $30/mo cap the standard DeepSeek V4 Flash carries on the same table.

      OpenCode Go: no DeepSeek V4.1 Flash row → OpenCode Go: DeepSeek V4.1 Flash at $0.15/$0.60/$0.003 off-peak (peak 2×), $15/mo usage cap

      Verified 2026-09-12 · source

    • ModelsCommandCode

      DeepSeek V4.1 Flash (launched 2026-09-10) is available on every CommandCode plan — the published availability map sets Go/GOAT/Pro/Provider/Max/Ultra/Teams all true. Per-model monthly credits $40 on GOAT ($50 on Pro), temporarily boosted to $60/$70 through 2026-09-17; per-token rates unchanged.

      CommandCode catalog: no DeepSeek V4.1 Flash row → CommandCode: DeepSeek V4.1 Flash on every plan (Go and above) at $0.15/$0.60/$0.003 off-peak (peak 2×); $40/mo GOAT cap boosted to $60 through 2026-09-17

      Verified 2026-09-12 · source

    • LaunchStandard Compute

      Standard Compute added to the catalogue: an OpenAI-compatible, smart-routed flat-rate API (base https://api.stdcmpt.com/v1, model id "standardcompute") whose plans sell a dollar-denominated monthly compute budget spent at provider list rates, then HTTP 402 until renewal.

      Starter $19 / Economy $39 / Standard $89 / Pro $249 (individual) and Pro Plus $499 / Growth $999 / Scale $2,499 (business) — flat monthly price with a fixed dollar compute budget of $20/$41/$95/$269 and $549/$1,119/$2,849, the same 34-model smart-routed catalogue on every plan, no per-token meter and no 5h/week windows; 1.5x compute budget in month one on individual plans

      Verified 2026-09-12 · source

    • LaunchZenMux

      ZenMux added to the catalogue: a Singapore-operated AI gateway whose Builder Plan sells a fixed monthly fee for a Flow quota (a currency-like usage unit accounting for tokens plus call overhead) on any OpenAI/Anthropic-compatible agent; premium model coverage rotates on the lower tiers.

      Builder Starter $20 / Builder Max $100 / Builder Ultra $200 — Flow-metered quotas of 50 / 300 / 800 Flows per rolling 5-hour window (1 Flow ≈ $0.03283), 10-15 RPM, rolling weekly limit, 4 / 3 / 2 manual 5-hour window resets per month, one OpenAI/Anthropic-compatible key for a 100+ model catalogue; non-production use only

      Verified 2026-09-12 · source

  14. 11 Sept 2026

    • ModelsLLM Gateway DevPass

      Two new Sakana models joined the DevPass coding catalogue on 2026-09-11 — Fugu Max ($2/$6 per 1M) and Fugu Ultra v2.0 ($5/$30 per 1M, premium tier). The Lite / Pro / Max headline arrays keep their curated flagship set, so only the catalogue count moves (123 → 125).

      DevPass coding catalogue: 123 models (2026-09-10) → DevPass coding catalogue: 125 models (2026-09-11) — Fugu Max and Fugu Ultra v2.0 added; headline arrays unchanged

      Verified 2026-09-12 · source

    • ModelsNeural Watt

      Neural Watt pricing page grew to 14 models; GLM 5.3 ($1.45/$4.50, 1M context, exited preview 2026-09-10) and Qwen 3.8 27B ($0.45/$3.20) are new on Basic/Standard/Pro. Kimi K3 has been on the pricing table since integration (present in every dump since 2026-08-09) but was missing from stored plan arrays — first stored 2026-09-11 as a site correction.

      5 models on Basic / Standard / Pro plans → 8 models on Basic / Standard / Pro plans — GLM 5.3 and Qwen 3.8 27B added, Kimi K3 stored-array gap filled

      Verified 2026-09-11 · source

  15. 10 Sept 2026

    • ModelsLLM Gateway DevPass

      LLM Gateway DevPass coding catalogue grew from 110 to 123 models between the Aug-21 and Sep-10 dumps. GPT-6 Astra (OpenAI, $10/$50/$1.00 per 1M) and Muse Spark 1.3 (Meta, $1.25/$4.25/$0.15 per 1M) added to the DevPass Lite / Pro / Max headline model arrays.

      DevPass coding catalogue: 110 models (2026-08-21) → DevPass coding catalogue: 123 models (2026-09-10) — GPT-6 Astra and Muse Spark 1.3 added to Lite / Pro / Max headline arrays

      Verified 2026-09-10 · source

    • PriceLLM Gateway DevPass

      Site correction: Nemotron 3 Ultra 550B output rate on DevPass Lite / Pro / Max was stored at $2.50/M, which never matched the plan's own source. Corrected to $2.20/M per the deepinfra row on llmgateway.io/models (2026-08-21 and 2026-09-10 dumps both show $0.50 in / $2.20 out).

      Nemotron 3 Ultra 550B output stored at $2.50/M → Nemotron 3 Ultra 550B output corrected to $2.20/M (deepinfra row; verified 2026-09-10)

      Verified 2026-09-10 · source

    • ModelsOpenAI

      OpenAI added GPT-6 Astra to the Codex usage table: Plus 5-45 messages per window, Pro 5× 25-225 and Pro 20× 100-900. Its FAQ states Astra is included within the existing Codex allowance on the $100/$200 Pro plans and that Plus gets limited Astra usage. It is not offered on the $8 Go tier, which stays without it.

      GPT-6 Astra not listed on any tracked OpenAI Codex plan → GPT-6 Astra added to OpenAI Plus, Pro 5× and Pro 20× (lead flagship on each)

      Verified 2026-09-10 · source

    • ModelsCommandCode

      CommandCode added OpenAI GPT-6 Astra at $10/$50 per 1M (cache read $1.00, cache write $12.50). Its published availability map enables the model for Max 10×, Max 20× and Provider only, so Go, GOAT and Pro stay without it.

      GPT-6 Astra not listed on any CommandCode plan → GPT-6 Astra added to Max 10×, Max 20× and Provider (Go, GOAT and Pro excluded)

      Verified 2026-09-10 · source

    • ModelsNous Portal

      Nous Portal added OpenAI GPT-6 Astra to the subscription model list at $8/$40 per 1M input/output with $0.80 cached read — below OpenAI list. Super and Ultra scale the Plus budget 5×/10×, so all three tiers gain it.

      GPT-6 Astra not in the Nous Portal subscription model list → GPT-6 Astra added to Nous Plus, Super and Ultra (openai/gpt-6-astra at $8/$40 per 1M)

      Verified 2026-09-10 · source

  16. 08 Sept 2026

    • PriceNous Portal

      Nous Portal inference list prices moved across all three Nous plans: DeepSeek V4 Pro output halved ($1.584 → $1.392) while input and cached rose; DeepSeek V4 Flash and Kimi K2.6 moved modestly.

      DeepSeek V4 Pro: $0.528/$1.584/$0.0176; DeepSeek V4 Flash: $0.0594/$0.1187/$0.0119; Kimi K2.6: $0.4484/$1.888/$0.0755 (in/out/cached $/1M) → DeepSeek V4 Pro: $0.696/$1.392/$0.1392; DeepSeek V4 Flash: $0.0543/$0.1344/$0.0134; Kimi K2.6: $0.4636/$1.952/$0.0781 (in/out/cached $/1M)

      Verified 2026-09-08 · source

    • ModelsNous Portal

      Flagship refresh across Nous Plus, Super, and Ultra: head chips updated to Claude Opus 5 (anthropic/claude-opus-5), GPT-5.6 Sol (openai/gpt-5.6-sol), and Gemini 3.7 Flash (google/gemini-3.7-flash); prior flagships retained in the granted list.

      Nous Plus/Super/Ultra flagship head chips: Claude Opus 4.7 / GPT-5.5 / Gemini 3.1 Pro → Nous Plus/Super/Ultra flagship head chips: Claude Opus 5 / GPT-5.6 Sol / Gemini 3.7 Flash

      Verified 2026-09-08 · source

  17. 04 Sept 2026

    • PriceCursor

      Data-integrity correction (config mismatch): Cursor exposes GPT-5.6 Sol at the OpenAI xHigh config ($4 input / $5 cache-write / $0.4 cache-read / $20 output), not OpenAI's base-promo list $2/$10/$0.20. The stored Cursor rule was set to the base-promo proxy; corrected to Cursor's own listed row (verified live 2026-09-04). Sol now renders ~2× more expensive on cursor-pro/pro-plus/ultra.

      GPT-5.6 Sol per-token rate on Cursor plans stored at $2/$10/$0.20 (OpenAI base-promo list) → GPT-5.6 Sol per-token rate on Cursor plans corrected to $4/$20/$0.4 (cached write $5) — cursor.com/docs/models-and-pricing own row

      Verified 2026-09-04 · source

    • ModelsCommandCode

      CommandCode pricing-limits availability matrix re-read, logged as one event for two model additions: Claude Fable 5.1 on the Max 10x / Max 20x / Provider tiers (individual-provider/max/ultra = true, pro = false, so Pro keeps Claude Fable 5), and Gemini 3.8 Flash on GOAT and above at list $1.50/$7.50/$0.15 (marked "Available on GOAT and above"). Gemini 3.7 Flash -50% (through Dec 31) remains a separate act-alone deal.

      Claude Fable 5.1 and Gemini 3.8 Flash not listed → Claude Fable 5.1 added to Max 10x / Max 20x / Provider (alongside Fable 5; Pro unchanged, keeps Fable 5); Gemini 3.8 Flash added to GOAT and above (Go unchanged) at list $1.50/$7.50/$0.15

      Verified 2026-09-04 · source

    • ModelsCursor

      Cursor model matrix re-read, logged as one event for two additions to the Pro, Pro+ and Ultra plan model lists: Claude Fable 5.1 at the $10/$50/$0.25 cache-read/$12.5 cache-write row alongside the retained Fable 5 row, and Gemini 3.8 Flash at the $0.75/$3.5/$0.075 model-pricing row alongside the retained Gemini 3.7 Flash row.

      Claude Fable 5.1 and Gemini 3.8 Flash not listed on Pro / Pro+ / Ultra → Claude Fable 5.1 ($10/$50/$0.25 cache-read/$12.5 cache-write) and Gemini 3.8 Flash ($0.75/$3.5/$0.075) added to Pro / Pro+ / Ultra

      Verified 2026-09-04 · source

    • ModelsGitHub Copilot

      GitHub Copilot supported-models catalog lists Claude Fable 5.1; added to the Pro+ and Max plan model arrays alongside the existing Fable 5 entry (credits/fair-use — no per-token math).

      Claude Fable 5 on Pro+ and Max plans → Claude Fable 5.1 added to Pro+ and Max plans (alongside Fable 5)

      Verified 2026-09-04 · source

    • ModelsLLM Gateway DevPass

      LLM Gateway model pages for claude-fable-5-1 and gemini-3.8-flash are both live (status active), logged as one event: Fable 5.1 ($10/$50/$0.25/$12.5) joins the DevPass Lite / Pro / Max model arrays behind the Claude Opus 5 flagship, and Gemini 3.8 Flash joins them at the $0.75 input / $3.75 output / $0.075 cache-read row.

      Claude Fable 5.1 and Gemini 3.8 Flash not listed on DevPass Lite / Pro / Max → Claude Fable 5.1 ($10/$50/$0.25/$12.5) and Gemini 3.8 Flash ($0.75 input / $3.75 output / $0.075 cache-read) added to DevPass Lite / Pro / Max

      Verified 2026-09-04 · source

  18. 03 Sept 2026

    • ModelsGitHub Copilot

      GitHub Copilot Pro/Pro+/Max flagship chips refreshed from the live 28-model catalog matrix. Pro: GPT-5.5 → GPT-5.6 Terra, Gemini 3.5 Flash → 3.7 Flash as headline chips (GPT-5.5 and GPT-5.6 Sol are strikethrough/excluded on Pro). Pro+ and Max: Claude Opus 5 + GPT-5.6 Sol + Gemini 3.7 Flash as new first-3 chips; Opus 4.8 retained in the curated array below Opus 5. Kimi K2.7 Code and Gemini 3.5 Flash remain in all three arrays but demoted below Kimi K3 and Gemini 3.6/3.7 Flash respectively.

      Pro flagships: Claude Sonnet 5, GPT-5.5, Gemini 3.5 Flash, Kimi K2.7 Code; Pro+ flagships: Claude Opus 4.8, GPT-5.5, Gemini 3.5 Flash, Kimi K2.7 Code; Max flagships: Claude Opus 5, GPT-5.6 Sol, Gemini 3.5 Flash, Kimi K2.7 Code → Pro flagships: Claude Sonnet 5, GPT-5.6 Terra, Gemini 3.7 Flash (+10 curated canonical models); Pro+ and Max flagships: Claude Opus 5, GPT-5.6 Sol, Gemini 3.7 Flash (+18 curated canonical models). Catalog refreshed against live 28-model matrix — GPT-5.5 and GPT-5.6 Sol excluded on Pro (strikethrough); Gemini 3.5 Flash demoted below 3.7/3.6; Claude Opus 5 now on Pro+ alongside Opus 4.8.

      Verified 2026-09-03 · source

    • ModelsGoogle AI

      Google launches Gemini 3.8 Flash, its newest Flash-tier flagship for coding and agents (released September 2026), as the follow-on to Gemini 3.7 Flash. Intro API rate is $0.75/$3.75 (input/output) through 2026-12-31, then $1.50/$7.50. Ranked #15 (High) on the LMArena agent leaderboard.

      Gemini 3.8 Flash — Google newest Flash flagship for coding and agents, intro $0.75/$3.75 (input/output) through 2026-12-31 then $1.50/$7.50

      Verified 2026-09-04 · source

  19. 02 Sept 2026

    • PriceYolo-Auto

      Yolo-Auto re-tiers pricing: Starter ($6) discontinued; Builder $10 → $19 (~3-4 coding agents, 128K context) and Pro $15 → $39 (~5-6 coding agents, 256K context); adds a Free $0 fair-use trial tier (not tracked); paid tiers reframed from "concurrent units" to "coding agents".

      Starter/Builder/Pro $6/$10/$15 → Free (trial) / Builder $19 / Pro $39 — Starter discontinued

      Verified 2026-09-02 · source

    • CeilingOllama

      Ollama Cloud moved from multipliers to a credits model and unpaused Max. Pro ($20) and Max ($100) now include month of usage credits ($60 / $300) at published per-token rates (replacing the community-derived proxy cost rules with exact rates), access to larger pro models, and early access to newest models on Max. Usage refreshes monthly with no rollover; extra usage is pay-as-you-go. Team ($500/mo → $1,000 credits) and Enterprise (Custom) exist but are not tracked.

      Pro $20/mo (50× Free usage multiplier), Max $100/mo (5× Pro usage) — per-GPU/usage-level metering, no published per-token math; Max sign-ups paused → Pro $20/mo → $60 usage credits, Max $100/mo → $300 usage credits (monthly, no rollover, pay-as-you-go overage); Max unpaused (10 concurrent requests); Ollama now publishes per-token rates for all cloud models

      Verified 2026-09-02 · source

    • LaunchMuse Code

      Meta Muse Code launches 3 fixed monthly subscription plans (out of beta Aug 31 2026), model = Muse Spark 1.2, exclusive to the Muse Code CLI: Everyday $5/mo (10–50 requests/5h; voice, web search, image uploads), High $15/mo (3× Everyday usage, latest models), Power $50/mo (10× Everyday usage, early access). Prices per TheNewStack launch report 2026-09-01 (community-corroborated $5/$15/$50; supersedes an earlier $18/$72/$160 community report) — reconcile against dev.meta.ai onboarding once publicly priced.

      Verified 2026-09-02 · source

    • ModelsSynthetic· Subscription

      Synthetic.new Subscription ($30/mo) swaps Qwen3.6-27B for Qwen3.8-27B at unchanged rates and adds GLM-5.3-Flash (Beta).

      6 models — Kimi K3, GLM-5.2, Nemotron 3 Super 120B, Qwen3.6-27B, GPT-OSS 120B, GLM-4.7-Flash → 7 models — Qwen3.6-27B → Qwen3.8-27B ($0.45/$2.20/$0.09 unchanged); GLM-5.3-Flash added ($0.15/$0.50/$0.04, Beta)

      Verified 2026-09-02 · source

  20. 01 Sept 2026

    • ModelsClaude

      Anthropic launches Claude Fable 5.1 (released 2026-09-01), the successor to Claude Fable 5 as its next-generation agentic-coding flagship. It keeps the $10/$50 (input/output) API rate of Fable 5 and cuts cache-read pricing to $0.25/MTok. New registry + model page here; per-provider plan availability tracked by the sibling provider bead.

      Claude Fable 5.1 — Anthropic’s next-generation agentic-coding flagship, same $10/$50 (input/output) API rate as Fable 5 with cache reads cut to $0.25/MTok

      Verified 2026-09-04 · source

    • ModelsOpenAI

      OpenAI launches GPT-6 Astra (API gpt-6-astra), its frontier model above the GPT-5.6 line (released 2026-09-01) at $10/$50 (input/output, cached input $1, cache write $12.50), rolling out to ChatGPT Plus/Pro/Business/Enterprise and AWS/Bedrock in the coming days. No tracked provider lists it as of 2026-09-04.

      GPT-6 Astra (API gpt-6-astra) — OpenAI frontier model above the GPT-5.6 line at $10/$50 (input/output, cached input $1, cache write $12.50)

      Verified 2026-09-04 · source

  21. 31 Aug 2026

    • PriceOpenAI· Plus

      Data-integrity correction (copy/paste): GPT-5.6 Sol was stored at GPT-5.5's $5/$30/$0.50 in the OpenAI plans, Cursor, CommandCode, LLM Gateway DevPass and Copilot Max per-model rules. Verified against platform.openai.com/docs/pricing — GPT-5.6 Sol list is $2/$10/$0.20 (cached write $2.50), promo to Nov 21 2026. Sol now renders ~2.5× cheaper than before on those plans.

      GPT-5.6 Sol per-token rate stored at $5/$30/$0.50 → GPT-5.6 Sol per-token rate corrected to $2/$10/$0.20 (OpenAI list; promotional pricing through Nov 21 2026)

      Verified 2026-08-31 · source

    • PriceOzore

      Ozore restructured from a 4-plan premium-token-pool model to a 2-plan credits model. Starter and Max discontinued; Basic and Pro bill monthly credits ($20 / $70) spendable on any model at per-token discounted rates (35% off maker list). Pro adds Smart Compression, Analytics Pro, Privacy Pack, 10 extra API keys, Custom Router, and Fusion. Hero lineup: GPT-5.6 Sol, Claude Opus 5, Gemini 3.7 Flash, GLM-5.3, DeepSeek V4 Pro, MiniMax M3, MiMo V2.5 Pro, Grok 4.6, GLM-5.3 Flash, HY3 and Nemotron 3 Ultra free.

      4 plans — Starter $9 / Basic $18 / Pro $32 / Max $50 with premium token pools (6M/13M/25M/40M premium tokens/mo + direct-pin daily pools) → 2 plans — Basic $10 ($20 credits/mo) / Pro $35 ($70 credits/mo + power features); credits billing at per-token discounted rates

      Verified 2026-08-31 · source

    • PriceCline· Pass

      Site correction: ClinePass bills as a flat-fee quota meter, so these are the plan’s per-token REFERENCE rates used by the calculator. DeepSeek V4 Pro/Flash were storing the neighboring MiMo V2.5 Pro / MiMo V2.5 rows (copy/paste); corrected to the live off-peak rates on docs.cline.bot.

      DeepSeek V4 Pro $1.74/$3.48, DeepSeek V4 Flash $0.14/$0.28 (per-1M reference; copy/paste of MiMo rows) → DeepSeek V4 Pro $0.66/$1.98, DeepSeek V4 Flash $0.22/$0.66 (off-peak reference, verified live 2026-08-31)

      Verified 2026-08-31 · source

    • ModelsZ.ai

      Z.AI GLM Coding Lite/Pro/Max supported models updated: GLM-5.3-Flash (multimodal flash tier, 2.3/0.56/8 multipliers) added; GLM-5.3-Flash replaces GLM-5-Turbo; GLM-4.7 requests now auto-route to GLM-5.3-Flash.

      GLM-5.3, GLM-5-Turbo, GLM-4.7 callable on coding plans (GLM-4.7 routed to GLM-5.3) → GLM-5.3 + GLM-5.3-Flash callable (GLM-5.3-Flash replaces GLM-5-Turbo; GLM-4.7 routes to GLM-5.3-Flash)

      Verified 2026-08-31 · source

    • ModelsCommandCode

      CommandCode adds five open models to all six individual plans: Qwen3.8 Flash, Qwen 3.8 27B, Hy4 preview, DeepSeek V4 Flash Fast, and DeepSeek V4 Flash Vision Exp.

      Go/GOAT/Pro/Max 10×/Max 20×/Provider model arrays previously synced 2026-08-17 (GOAT 2026-08-28) → Qwen3.8 Flash, Qwen 3.8 27B, Hy4 preview, DeepSeek V4 Flash Fast, DeepSeek V4 Flash Vision Exp added to all six plans

      Verified 2026-08-31 · source

  22. 30 Aug 2026

    • ModelsOpenCode· Go

      OpenCode Go adds four models: LongCat-2.0 (Meituan 1.6T MoE), Qwen3.8 Flash, DeepSeek V4 Flash Vision Exp (vision input, $15/mo cap — half the standard V4 Flash cap), and Hy4 preview (Tencent).

      Go model lineup lacked LongCat-2.0, Qwen3.8 Flash, DeepSeek V4 Flash Vision Exp, Hy4 preview → Go model lineup adds LongCat-2.0 ($0.30/$1.20, $60/mo), Qwen3.8 Flash ($0.15/$0.47, $30/mo), DeepSeek V4 Flash Vision Exp ($0.22/$0.66, $15/mo, peak 2×), Hy4 preview ($0.834/$2.501, $30/mo)

      Verified 2026-08-30 · source

  23. 29 Aug 2026

    • ModelsCursor

      Cursor refreshes the Other Models pool on Pro, Pro+ and Ultra: adds Gemini 3.7 Flash and GPT-5.6 Terra, and withdraws GLM-5.2, Kimi K3, Kimi K2.7 Code, GPT-5.5 and GPT-5.4.

      Other Models pool: GLM-5.2, Kimi K3, Kimi K2.7 Code, GPT-5.5, GPT-5.4 offered; Gemini 3.7 Flash and GPT-5.6 Terra not offered → Other Models pool: Gemini 3.7 Flash + GPT-5.6 Terra added; GLM-5.2, Kimi K3, Kimi K2.7 Code, GPT-5.5, GPT-5.4 withdrawn

      Verified 2026-08-29 · source

    • PolicyCursor

      OpenAI plans to end its partnership with Cursor following SpaceX acquisition (Anysphere), proposing a shutoff date of 2026-11-12 and citing concerns about terms compliance. Reuters-confirmed 2026-08-28. Current GPT-5.6 Luna/Sol/Terra availability in Cursor is unchanged until then.

      OpenAI models (GPT-5.6 Luna/Sol/Terra) available in Cursor via OpenAI partnership → OpenAI to end Cursor partnership — OpenAI model access in Cursor expected to end 2026-11-12; no future OpenAI models provided (forward-looking, no change to current availability)

      Verified 2026-08-29 · source

    • LaunchYolo-Auto

      Yolo-Auto launches 3 flat-rate, unlimited-token OpenAI-compatible plans serving Qwen3.8-27B (text + vision input) with no per-token meter. Starter $6 (1 concurrent unit, 128K context), Builder $10 (2 units), Pro $15 (4 units, priority on shared capacity). Base URL https://yolo-auto.com/v1; works with any OpenAI-compatible client (Hermes, Cursor, Claude Code via LiteLLM, pi, openclaw). No per-token or cap ceilings published.

      Yolo-Auto flat-rate coding plans — Starter / Builder / Pro ($6/$10/$15/mo)

      Verified 2026-08-29 · source

  24. 28 Aug 2026

    • PolicyOpenCode· Go

      OpenCode Go ended the $5 first-month promo; the plan is $10/mo flat from the first month.

      $5 first month / 50% off 1st mo → no first-month discount ($10/mo from month one)

      Verified 2026-08-28 · source

    • ModelsOpenCode· Go

      GLM-5.3-Flash is now listed on OpenCode Go at $0.15/$0.50/$0.03 per 1M tokens with a $15 monthly usage cap.

      GLM-5.3-Flash not listed → GLM-5.3-Flash added (Z.AI, $0.15/$0.50/$0.03, $15 cap)

      Verified 2026-08-28 · source

    • ModelsCommandCode

      CommandCode GOAT now serves GPT-5.6 Sol (previously excluded from the GOAT model set and only available on higher tiers); per commandcode.ai/docs/resources/pricing-limits the model is available on GOAT and above at $5/$30/$0.5 with a $70 GOAT credit allowance.

      GOAT plan did not offer GPT-5.6 Sol → GOAT plan now offers GPT-5.6 Sol

      Verified 2026-08-28 · source

    • PolicyZ.ai

      Site correction: GLM Coding plans are restricted to an official allowlist of supported tools, not any tool. Z.AI docs state users may not use subscription benefits for tools or scenarios outside the supported scope.

      GLM Coding Lite/Pro/Max usable with any tool · API key (codingTool "Any tool") → GLM Coding Lite/Pro/Max limited to officially supported tools (ZCode, Claude Code, Cline, OpenCode, Cursor, Codex, Kilo Code, Roo Code, Crush, Goose, etc.)

      Verified 2026-08-28 · source

  25. 26 Aug 2026

    • ModelsOpenCode· Go

      OpenCode Go adds Grok 4.6 to its model lineup and drops Grok 4.5 (docs now list only Grok 4.6: model id grok-4.6, ~169 req/5h / 845/mo, $15/mo usage, 30-day retention).

      Grok 4.5 only → Grok 4.6 only (replaces Grok 4.5; $15/mo usage, 30-day retention)

      Verified 2026-08-26 · source

    • LaunchPhoenix Grove

      Phoenix Grove Systems launches 6 OpenAI-compatible coding plans (any SDK/agent/tool via api.pgsgrove.com/v1). Taster $3.99 (first month free, DeepSeek V4 Flash only), Basic $12.95, Pro $25, Elite $50, Ultra $99, Canopy $195; standard-monthly + 8× light-model token budgets with 3–8 agent concurrency.

      Phoenix Grove (PGS) coding plans — Taster / Basic / Pro / Elite / Ultra / Canopy ($3.99–$195/mo)

      Verified 2026-08-26 · source

    • ModelsCommandCode

      Identity correction: the CommandCode model previously cataloged as "Ox Alpha" is confirmed as GLM-5.3-Flash (family zai), a first-party Z.AI model. Z.AI’s own pricing page lists GLM-5.3-Flash as a Latest Model (docs.z.ai/guides/overview/pricing) at $0.15 input / $0.075 cached / $0.50 output, with a 50% launch discount through 2026-09-09. Z.AI confirmed the identity to Bloomberg on 2026-08-26; weights open that night (MIT). GLM-5.3-Flash is a separate entry from the GLM-5.3 flagship; the stealth/ox-alpha route id on OpenRouter/CommandCode is unchanged. Still billed $0/M during its CommandCode preview window on Go and above, distinct from Z.AI’s API list price.

      Model cataloged as "Ox Alpha" (family: stealth, anonymous lab) → GLM-5.3-Flash, family zai — reidentified as a confirmed Z.AI model. Listed in Z.AI’s "Latest Models" pricing (docs.z.ai/guides/overview/pricing) at $0.15/$0.075/$0.50, 50% promo through 2026-09-09. Confirmed via Bloomberg 2026-08-26; weights open, MIT.

      Verified 2026-08-26 · source

    • CeilingOpenAI· Plus

      Luna Reserve gives selected ChatGPT Plus accounts additional GPT-5.6 Luna-only usage after their regular Codex/ChatGPT Work usage is exhausted; it has its own separate usage limit and is a fallback mode, not a separate model to buy. Reddit r/codex claim (1vyzqgx / 1vz8l4i, 2026-08-26) confirmed against official help.openai.com.

      No Luna fallback reserve → Luna Reserve: extra Luna-only usage after regular weekly cap exhausted (own limit, selected Plus accounts)

      Verified 2026-08-28 · source

  26. 25 Aug 2026

    • ModelsX.AI· SuperGrok Lite

      xAI adds Grok 4.6 to SuperGrok Lite ($10). All four SuperGrok tiers (Lite $10, SuperGrok $30, Plus $100, Heavy $300) now include both Grok 4.5 and Grok 4.6.

      Grok 4.5 only → Grok 4.5 + Grok 4.6

      Verified 2026-08-25 · source

  27. 24 Aug 2026

    • LaunchStepFun

      StepFun Step Plan added (credit-based subscription, $1≈7M credits): Flash Mini $6.99/400M · Flash Plus $9.99/1600M · Flash Pro $29/8000M · Flash Max $99/40000M credits per month (cleared at month end, no rollover); Step 3.7 Flash + Step 3.5 Flash; BYOK OpenAI-compatible endpoint api.stepfun.ai/step_plan/v1 (Claude Code, Cursor, Cline, etc.).

      Verified 2026-08-24 · source

    • PriceCursor

      Claude Sonnet 5 API pricing: the $2/$10 launch intro price is now the permanent standard — Anthropic cancelled the previously scheduled $3/$15 increase from Sept 1. tokenplans refreshed the Cursor Other Models rate rows accordingly.

      Claude Sonnet 5 $3/$15 (list) on the Cursor Other Models pool → Claude Sonnet 5 $2/$10 — Anthropic made the launch intro price permanent (no Sept-1 increase)

      Verified 2026-08-24 · source

  28. 22 Aug 2026

    • ModelsOpenCode· Go

      OpenCode Go adds MiniMax M2.5 (relisted after an Aug 8 removal) and Muse Spark 1.2 Contributor (limited regions) to its model lineup.

      MiniMax M2.5 and Muse Spark 1.2 Contributor not included → MiniMax M2.5 and Muse Spark 1.2 Contributor included

      Verified 2026-08-22 · source

  29. 21 Aug 2026

    • ModelsLLM Gateway DevPass

      LLM Gateway DevPass Lite/Pro/Max flagship models refreshed to Claude Opus 5, Gemini 3.7 Flash and GPT-5.6 Sol, with Grok 4.6, Qwen3.8-Max and GLM-5.3 added to the headline lineup.

      Claude Opus 4.8, GPT-5.5, Gemini 3.1 Pro flagships; Kimi K3, Muse Spark 1.1, GLM-5.2, DeepSeek V4 Pro, MiMo V2.5 Pro, MiniMax M3, Nemotron 3 Ultra 550B, Qwen3.7-Flash → Claude Opus 5, Gemini 3.7 Flash, GPT-5.6 Sol flagships; Grok 4.6, Qwen3.8-Max, GLM-5.3, Kimi K3, MiniMax M3, Nemotron 3 Ultra 550B, Muse Spark 1.2, DeepSeek V4 Pro

      Verified 2026-08-21 · source

    • ModelsCommandCode

      CommandCode adds Ox Alpha, a stealth reasoning model from an anonymous provider, to all plans Go and above. Free during the preview window ($0/M input, output, cache reads). OpenCode also announced availability on X (Aug 20) but docs do not yet list it — left out of opencode-go plan models pending verification.

      No Ox Alpha → Ox Alpha (stealth/ox-alpha) — stealth frontier model from anonymous lab, 1M context, multimodal input (text+image+video), free during preview window (~1 week from Aug 20). Available on Go and above (Go, GOAT, Pro, Provider, Max 10×, Max 20×).

      Verified 2026-08-21 · source

    • PolicyCursor· Pro+

      Grok Bot (always-on coding agents) is now bundled with Cursor Pro+ ($60) — a dedicated weekly usage pool separate from the Cursor Models / Other Models pools. Included on Pro+ and Ultra, NOT on Pro ($20) per official Cursor docs (help/grok-bot/plans: "Pro does not include Grok Bot access"; pricing page lists "Access to Grok Bot" on Pro+/Ultra only). Reddit r/cursor claims of it on Pro and a "2B Grok 4.6 tokens" count are not published officially — not recorded.

      No Grok Bot access → Grok Bot bundled (dedicated weekly usage pool, separate from Cursor model pools)

      Verified 2026-08-28 · source

  30. 18 Aug 2026

    • PriceCommandCode

      CommandCode per-token promos ended: GPT-5.6 Luna and GPT-5.6 Terra revert from -50% deal-applied to list rates (2x usage cost on those models across Go, GOAT, Pro, Max 10x, Max 20x and Provider), and DeepSeek V4 Pro / V4 Flash move from -75% deal rates to time-of-day pricing (off-peak headline; peak 2x, 7h/day) — roughly 1.5–2.3x the previous usage cost on DeepSeek rows. Gemini 3.7 Flash -50% (through Dec 31 2026) is separate and still active.

      Deal-applied per-token rates: GPT-5.6 Luna $0.10/$0.60 (cached $0.01), GPT-5.6 Terra $1.00/$6.00 (cached $0.10) (-50% promos through 2026-08-13); DeepSeek V4 Pro $0.435/$0.87, DeepSeek V4 Flash $0.14/$0.28 (-75% promos) → List rates from 2026-08-18 live docs: GPT-5.6 Luna $0.20/$1.20 (cached $0.02), GPT-5.6 Terra $2.00/$12.00 (cached $0.20); DeepSeek V4 Pro $0.66/$1.98 (cached $0.022) and V4 Flash $0.22/$0.66 (cached $0.007) at off-peak headline, peak 2x for 7h/day (01–04 & 06–10 UTC) since 2026-08-16

      Verified 2026-08-18 · source

  31. 17 Aug 2026

    • PolicyCommandCode

      Site correction surfacing CommandCode API access — every plan except the $1 Go plan can call the Provider API (OpenAI/Anthropic-compatible endpoints, no lock-in); coding plans meter API usage against their own credits and per-model access.

      API access on the Provider plan only; GOAT, Pro, Max 10x and Max 20x listed as CLI-exclusive → API access on every plan except Go — GOAT, Pro, Max 10x, Max 20x and Provider serve OpenAI- and Anthropic-compatible endpoints usable by any coding tool; Go remains CLI-only (403 upgrade_required)

      Verified 2026-08-17 · source

    • CeilingOpenCode· Go

      OpenCode Go raises the DeepSeek V4 Flash monthly usage cap back to $30/mo on the live docs (was $15 on Aug 16 after the 2× promo withdrawal); request allowance roughly doubles to ~37,800 req/mo. Published request shape is 410 input / 71,300 cached / 310 output.

      DeepSeek V4 Flash: $15 monthly usage budget (post-promo, from Aug 16) → DeepSeek V4 Flash: $30 monthly usage budget

      Verified 2026-08-17 · source

  32. 16 Aug 2026

    • CeilingOpenCode· Go

      OpenCode Go withdraws the DeepSeek V4 Flash 2× usage promo early (Aug 16) and caps both DeepSeek models at $15/mo usage — 4× below the $60 base; Flash request allowance cut ~8× to 18,900/mo.

      DeepSeek V4 Flash: $120 monthly usage budget (2× promo through Aug 31) → DeepSeek V4 Flash & V4 Pro: $15 monthly usage budget

      Verified 2026-08-16 · source

    • ModelsOpenCode· Go

      OpenCode Go adds GLM-5.3 to the Go model lineup.

      GLM-5.3 not offered → GLM-5.3 added ($15 usage cap)

      Verified 2026-08-16 · source

    • ModelsZ.ai

      Z.AI GLM Coding Lite/Pro/Max supported models updated to GLM-5.3 + GLM-5-Turbo + GLM-4.7; previous-model requests (GLM-5.2/5.1) are automatically routed to GLM-5.3 at the same credit multipliers.

      GLM-5.2 listed as flagship (GLM-5.2, GLM-5-Turbo, GLM-4.7) → GLM-5.3 flagship; GLM-5.2/5.1 requests auto-route to GLM-5.3 (GLM-5.3, GLM-5-Turbo, GLM-4.7)

      Verified 2026-08-16 · source

    • ModelsChutes

      Chutes adds DeepSeek V4 Flash to the public catalog and pricing page (TEE-hosted V4-Flash-0731 build, $0.14/$0.28 per 1M, 1.0M context). Both Plus and Pro model lists updated. DeepSeek V4 Pro is not listed anywhere on chutes.ai.

      Plus and Pro list only DeepSeek V3.2 (13 models) → Plus and Pro add DeepSeek V4 Flash (14 models; V4 Pro not offered)

      Verified 2026-08-16 · source

    • ModelsCommandCode

      CommandCode pricing-limits matrix re-read, logged as one event for two model additions: GLM-5.3 (Z.AI open-weight flagship, 1M context) on all six individual plans, and Gemini 3.7 Flash on GOAT and above (GOAT $40 / Pro $60 monthly allowance at deal-applied rates, effectively 2× usage via the -50% promo through Dec 31 2026) with Go unchanged.

      GLM-5.3 and Gemini 3.7 Flash not listed on any individual plan → GLM-5.3 added to Go/GOAT/Pro/Max 10×/Max 20×/Provider (1M context, reasoning); Gemini 3.7 Flash added to GOAT/Pro/Max 10×/Max 20×/Provider (1M context, -50% deal through Dec 31 2026)

      Verified 2026-08-17 · source

  33. 13 Aug 2026

    • ModelsNous Portal

      Nous Portal adds Grok 4.6 to the included model list of Plus, Super and Ultra (one change, three plans — single entry per aggregation rule).

      22 models per plan (Claude Opus 4.7 through Hy3) → 23 models per plan — Grok 4.6 added to Plus, Super and Ultra

      Verified 2026-08-13 · source

    • PriceNous Portal

      Nous Portal inference list prices moved for Kimi K2.6, GLM-5.1, DeepSeek V4 Pro, and Kimi K2.7 Code (-20% promo still applied; DeepSeek V4 Flash base slug unchanged at 0.112/0.224/0.0224).

      Kimi K2.6: $0.48/$2.73/$0.16; GLM-5.1: $0.77/$2.43/$0.14352; DeepSeek V4 Pro: $0.35/$0.70/$0.0029; Kimi K2.7 Code: $0.58/$2.80/$0.12 → Kimi K2.6: $0.76/$3.20/$0.128; GLM-5.1: $1.12/$3.52/$0.208; DeepSeek V4 Pro: $0.9344/$1.8688/$0.0788; Kimi K2.7 Code: $0.536/$2.72/$0.12

      Verified 2026-08-13 · source

  34. 12 Aug 2026

    • ModelsCline· Pass

      Cline Pass adds Qwen3.8 Max to its included model list.

      11 models (GLM-5.2 through Qwen3.7-Plus) → 12 models — Qwen3.8-Max added

      Verified 2026-08-12 · source

    • ModelsCommandCode

      CommandCode adds Grok 4.6 (xAI flagship, Aug 12 2026) to GOAT and above; the $1 Go plan keeps Grok 4.5 only (availability individual-go: false).

      Grok 4.5 only on GOAT and above (Grok 4.6 not offered; Go lists Grok 4.5) → Grok 4.6 added to GOAT, Pro, Provider, Max 10x and Max 20x; Go unchanged

      Verified 2026-08-16 · source

    • ModelsCursor

      Cursor adds Grok 4.6 (first-party Cursor Router model, 50% launch discount for one week from Aug 12) to the Cursor Models pool on Pro, Pro Plus and Ultra.

      Cursor Models pool: Grok 4.5 + Composer 2.5 → Cursor Models pool: Grok 4.6 + Grok 4.5 + Composer 2.5

      Verified 2026-08-16 · source

    • ModelsX.AI

      xAI adds Grok 4.6 to SuperGrok ($30), Plus ($100) and Heavy ($300); SuperGrok Lite keeps Grok 4.5.

      Grok 4.5 only on SuperGrok, Plus and Heavy (Lite unchanged) → Grok 4.6 added to SuperGrok, SuperGrok Plus and SuperGrok Heavy

      Verified 2026-08-16 · source

    • ModelsOzore

      Ozore adds Grok 4.6 to its premium token pool at 1.5x (between 1x Grok 4.3/4.5 and 2x Gemini 3.1 Pro) on all four plans.

      Grok 4.3 + Grok 4.5 + Grok Build 0.1 in premium pool → Grok 4.6 added to premium pool at 1.5x multiplier (667K tokens per 1M pool)

      Verified 2026-08-16 · source

  35. 11 Aug 2026

    • PolicyCommandCode

      CommandCode promo multiplier window active through Aug 13, 2026 — DeepSeek V4 Pro 4x, MiniMax M3 2x, MiMo up to 99% off, Laguna free, GPT-5.6 Luna up to 95% off, GPT-5.6 Terra 50% off.

      1× credit consumption (standard rates) → Multiplier window active through Aug 13, 2026: DeepSeek V4 Pro 4×, MiniMax M3 2×, MiMo up to 99% off, Laguna S 2.1 free, GPT-5.6 Luna up to 95% off, GPT-5.6 Terra 50% off

      Verified 2026-08-11 · source

    • ModelsOzore

      Ozore catalog expansion tracked — plan model lists updated from 19 to 47 models (25 premium + 22 standard); prices and pools unchanged.

      19 models tracked per plan → 47 models tracked per plan (25 premium + 22 standard)

      Verified 2026-08-11 · source

    • ModelsGoogle AI

      Site correction surfacing Google AI actual model lists — Gemini Spark (24/7 agent) available to Google AI Pro and Ultra subscribers; Gemini 3 Pro in AI Mode for Google Search available on Ultra tiers.

      Gemini Spark missing from Pro; Gemini 3 Pro missing from Ultra 5x/20x → Gemini Spark added to Pro; Gemini 3 Pro added to Ultra 5x and Ultra 20x

      Verified 2026-08-11 · source

    • PolicyXiaomi MiMo

      Xiaomi MiMo token plans now advertise unlimited usage with no weekly or 5-hour caps, plus a night 0.8x usage multiplier (00:00–08:00 UTC+8).

      Credit-total limits only (4.1B/11B/38B/82B credits monthly); no cap/policy wording published → Unlimited usage — no weekly cap, no 5-hour cap; night usage 0.8x (00:00–08:00 UTC+8)

      Verified 2026-08-11 · source

  36. 10 Aug 2026

    • PolicyQwenCloud

      QwenCloud cut the qwen3.8-max night discount from 80% extra to 50% off credits and temporarily lifted the 5-hour limit (7-day windows unchanged).

      Night discount 80% extra on qwen3.8-max; 5h limit 700/3,000/12,000 credits by plan → Night discount 50% off credits on qwen3.8-max (22:00–08:00 UTC+8); 5h limit temporarily lifted (unlimited)

      Verified 2026-08-10 · source

    • LaunchBytePlus

      BytePlus ModelArk Coding Plan added: Lite ($10/mo) and Pro ($50/mo, 5× Lite). Seed/GLM/Kimi/DeepSeek/GPT-OSS models, quota shared across tools.

      BytePlus ModelArk Coding Plan Lite ($10/mo, ~1,900 req/5h · ~12,000/wk · ~24,000/mo) and Pro ($50/mo, 5× Lite ≈ 9,500 req/5h · ~60,000/wk · ~120,000/mo, ArkClaw included)

      Verified 2026-08-10 · source

  37. 08 Aug 2026

    • CeilingOpenCode· Go

      OpenCode Go temporarily doubles DeepSeek V4 Flash usage.

      DeepSeek V4 Flash: $60 monthly usage budget → DeepSeek V4 Flash: $120 monthly usage budget

      Verified 2026-08-08 · source

    • ModelsOpenCode· Go

      OpenCode Go no longer serves MiniMax M2.5.

      MiniMax M2.5 included → MiniMax M2.5 removed

      Verified 2026-08-08 · source

    • CeilingZ.ai

      Z.AI GLM Coding limits confirmed as credit windows with dynamic 5-hour refresh and weekly reset.

      GLM Coding limits described as generic prompt quotas → 2,000/12,000/28,000 credits per 5h and 10,000/60,000/140,000 weekly credits; off-peak usage costs 50%

      Verified 2026-08-08 · source

    • PolicyKimi

      Kimi announced the Kimi Code separation, but its public pricing page still shows only paused membership tiers.

      Kimi and Kimi Code benefits presented together → Benefits will separate; no standalone Kimi Code plan published; existing membership tiers are waitlist-only

      Verified 2026-08-08 · source

  38. 02 Aug 2026

    • PolicyNous Portal

      Nous Portal runs limited-time inference promos — -20% all models (pinned Jul 22) and, since Aug 2, -50% on GPT-5.6 Terra/Luna and -90% on DeepSeek V4 Flash; the calculator stores deal-applied rates (Luna billed at -50% off list, DS4 Flash -90%).

      Nous Portal inference at list pricing (per API original block) → Promos: -20% across all models (pinned Jul 22, 2026); -50% GPT-5.6 Terra/Luna and -90% DeepSeek V4 Flash (posted Aug 2, 2026) — deal-applied rates stored in cost rules

      Verified 2026-08-08 · source

  39. 31 Jul 2026

    • PriceOpenAI

      OpenAI repriced GPT-5.6 Luna and Terra on 2026-07-31 — Luna list fell 80% ($1.00/$6.00 to $0.20/$1.20), Terra 20% ($2.50/$15.00 to $2/$12); Codex plan rows refreshed 2026-08-11 with Luna as the economy tier.

      GPT-5.6 Luna list $1.00/$6.00, GPT-5.6 Terra $2.50/$15.00 (pre-repricing baseline; per PRICING_SOURCES Nous section, API original block) → GPT-5.6 Luna list $0.20/$1.20 (-80%), GPT-5.6 Terra $2/$12 (-20%); Luna is now the economy Codex model — highest local-message limits (50-280/5h) vs Terra 20-110, Sol 15-90, and default on the Free tier

      Verified 2026-08-11 · source

    • PolicyOpenCode· Go

      OpenCode Go replaced the blanket zero-retention claim with a per-model data policy — the community-flagged ZDR wording removal (r/opencodeCLI 2026-08-01) is technically true of the wording, but the replacement is a more precise disclosure; retention 30d only for Grok 4.5 + GPT-5.6 Luna abuse logs.

      Blanket claim on Go docs: our providers follow a zero-retention policy (ZDR) → Per-model data policy: training Not used for all models; data retention 30 days for Grok 4.5 and GPT-5.6 Luna (abuse-monitoring logs), 0 days for every other model; ZDR remains explicitly in force for Grok 4.5 and DeepSeek V4 Flash (agreement renewed monthly, valid through Aug 31, 2026)

      Verified 2026-08-11 · source

Newsletter

Never miss a plan churn — get the weekly brief.

A plain weekly digest of every tracked price move, ceiling change and model-list update — pulled straight from the dated changelog ledger, with a one-line “worth switching?” note on each. Free. No spam, unsubscribe anytime.

Free · weekly · unsubscribe anytime