Every monthly plan that includes MiMo V2.6 Flash, converted to requests per $1 at the model's typical request shape. Same engine as the main calculator — proxy estimates carry a ~ marker; coarse estimates carry ≈.
MiMo V2.6 Flash (API mimo-v2.6-flash) is the low-cost tier of Xiaomi's three-model MiMo V2.6 series, released 2026-09-22 — full-modality reasoning built for high-frequency calls and large-scale workloads,.
Claude Fable 5.1Claude Opus 5.5Claude Opus 5GPT-6 AstraGPT-6 SolGPT-6 LunaGemini 3.8 FlashGemini 3.7 FlashGPT-5.6 SolGrok 4.6Grok 4.7Qwen3.8-MaxGLM-5.3Kimi K3MiniMax M3Nemotron 3 Ultra 550BMuse Spark 1.3Muse Spark 1.2MiMo V2.6 ProMiMo V2.6 FlashDeepSeek V4.1 FlashDeepSeek V4 Pro
Ceiling
$537/mo usage at provider rates · weekly frontier fair-use 18%
Reachability
Any agent · OpenAI-compatible key
Pricing notes
—
Cost rule
~proxy estimate · devpass.llmgateway.io/coding-models — mimo-v2.6-flash row (inputPrice 1.4e-7 / outputPrice 2.8e-7 per token = $0.14/$0.28 per 1M, premium: false, context 1M, recommended: true; read 2026-09-23). The $4/$20 reading taken when the sweep bead was filed was an upstream data-entry slip — the live row now matches Xiaomi list. Cached read from llmgateway.io/models — xiaomi mapping (same $0.14/$0.28, cachedInputPrice 2.8e-9 per token = $0.0028 per 1M; read 2026-09-23).
Claude Fable 5.1Claude Opus 5.5Claude Opus 5GPT-6 AstraGPT-6 SolGPT-6 LunaGemini 3.8 FlashGemini 3.7 FlashGPT-5.6 SolGrok 4.6Grok 4.7Qwen3.8-MaxGLM-5.3Kimi K3MiniMax M3Nemotron 3 Ultra 550BMuse Spark 1.3Muse Spark 1.2MiMo V2.6 ProMiMo V2.6 FlashDeepSeek V4.1 FlashDeepSeek V4 Pro
Ceiling
$87/mo usage at provider rates · weekly frontier fair-use 12%
Reachability
Any agent · OpenAI-compatible key
Pricing notes
—
Cost rule
~proxy estimate · devpass.llmgateway.io/coding-models — mimo-v2.6-flash row (inputPrice 1.4e-7 / outputPrice 2.8e-7 per token = $0.14/$0.28 per 1M, premium: false, context 1M, recommended: true; read 2026-09-23). The $4/$20 reading taken when the sweep bead was filed was an upstream data-entry slip — the live row now matches Xiaomi list. Cached read from llmgateway.io/models — xiaomi mapping (same $0.14/$0.28, cachedInputPrice 2.8e-9 per token = $0.0028 per 1M; read 2026-09-23).
Claude Fable 5.1Claude Opus 5.5Claude Opus 5GPT-6 AstraGPT-6 SolGPT-6 LunaGemini 3.8 FlashGemini 3.7 FlashGPT-5.6 SolGrok 4.6Grok 4.7Qwen3.8-MaxGLM-5.3Kimi K3MiniMax M3Nemotron 3 Ultra 550BMuse Spark 1.3Muse Spark 1.2MiMo V2.6 ProMiMo V2.6 FlashDeepSeek V4.1 FlashDeepSeek V4 Pro
Ceiling
$237/mo usage at provider rates · weekly frontier fair-use 15%
Reachability
Any agent · OpenAI-compatible key
Pricing notes
—
Cost rule
~proxy estimate · devpass.llmgateway.io/coding-models — mimo-v2.6-flash row (inputPrice 1.4e-7 / outputPrice 2.8e-7 per token = $0.14/$0.28 per 1M, premium: false, context 1M, recommended: true; read 2026-09-23). The $4/$20 reading taken when the sweep bead was filed was an upstream data-entry slip — the live row now matches Xiaomi list. Cached read from llmgateway.io/models — xiaomi mapping (same $0.14/$0.28, cachedInputPrice 2.8e-9 per token = $0.0028 per 1M; read 2026-09-23).
StepFunFlash Plus1,600M credits / month (no time-window or request-frequency caps)
$70 in credits/mo · + Smart Compression, Analytics Pro, Privacy Pack, 10 extra API keys, Custom Router, Fusion
Reachability
Any agent · OpenAI-compatible key
Pricing notes
Includes Custom Router, so you can pin a lineup and Ozore routes only among your picks. Cached-read pricing is not published, so the calculator estimates it at 10% of the input rate.
Cost rule
~proxy estimate · floor: cached reads bill at the input rate · ozore.com/api/credits/pricing (live 2026-09-23, billed row xiaomi/mimo-v2.6-flash $0.098/0.196 per 1M = 30% off list $0.14/0.28); cached-read not published → cached tokens billed at the input rate
OzoreBasic$20 in credits/mo · spend on any model at discounted rates · cancel anytime
$20 in credits/mo · spend on any model at discounted rates · cancel anytime
Reachability
Any agent · OpenAI-compatible key
Pricing notes
Requests are auto-routed — you do not pick the model, so a per-model figure is an upper bound on this tier; Direct Pin and Custom Router are Pro features. Cached-read pricing is not published, so the calculator estimates it at 10% of the input rate.
Cost rule
~proxy estimate · floor: cached reads bill at the input rate · ozore.com/api/credits/pricing (live 2026-09-23, billed row xiaomi/mimo-v2.6-flash $0.098/0.196 per 1M = 30% off list $0.14/0.28); cached-read not published → cached tokens billed at the input rate
Bills $3.99/mo after the free month · cancel anytime · entry tier (Flash-tier models) · one free trial month per person, duplicate trial accounts will be canceled
Unlimited tokens · 1 concurrent generation per stream (extra requests queue — queue waits unbounded, waits can be capped/dropped; 1–50 streams) · 260K context guaranteed
Reachability
Any agent · OpenAI/Anthropic-compatible key
Pricing notes
Meterless: no per-token rates are published, so there is no cost-rule basis. Model ID is "auto" — the model cannot be picked; the fleet rotates (live fleet listed here, the docs page is the authoritative list). The flat price is subsidized by a training/advertising licence over prompts and outputs (no opt-out).
60M input tokens/week · 2x on GLM 5/5.1/5.3, Kimi K2.7 Code, DeepSeek V4 Pro/Pro 0813 · 100 images/day
Reachability
Any agent · OpenAI-compatible key
Pricing notes
Not the best fit for coding agents — large prompts/context burn the weekly cap fast; PAYG safer for heavy coding. Web search is not included in PRO’s subscription coverage.
Price confirmed on the provider product page 2026-09-25 (dev.meta.ai/products/muse-code) · High Usage: 5× Everyday usage, more prompts with the latest Muse models
Cost rule
n/a
QwenCloudEssential Plan-37.5% off$10/mo5,625 Credits / 7 days
GPT-6 AstraGPT-6 SolGPT-6 LunaGPT-5.6 SolGPT-5.5GPT-5.6 TerraGPT-5.6 Luna
Ceiling
Unlimited* messages (reasonable use)
Reachability
—
Pricing notes
Codex limits (messages): GPT-6 Luna 350–3,000 = highest · GPT-6 Sol 15–150 · GPT-5.6 Luna 250–2,000 · GPT-5.6 Sol 10–100 · Terra 25–200 · Astra 5–45 · weekly limits may also apply · Luna Reserve = extra Luna-only usage once the regular weekly cap is exhausted (selected Plus accounts, own limit)
Basic models plus rotating limited-time premium models. Non-production use (personal dev / prototyping — PAYG is the production tier). Flow ≈ $0.03283 (floats live); per-model Flow burn and per-tier model lists are not published in text. Provider "Worth" claim: $30/mo equivalent API value (≈1.5× the fee), from its published equivalent-pay-as-you-go table.
Cost rule
n/a
QwenCloudStandard Plan-28% off$18/mo10,000 Credits / 7 days
Grok Bot access is included (not on Lite): link the Grok or X account to grant weekly Grok Bot usage on a Cursor account — cursor.com/help/grok-bot/plans
Cost rule
n/a
GitHub CopilotPro+$70/mo total AI credits · 4× Pro usage
Price confirmed on the provider product page 2026-09-25 (dev.meta.ai/products/muse-code) · Power Usage: 20× Everyday usage, early access to new features
GPT-6 AstraGPT-6 SolGPT-6 LunaGPT-5.6 Sol ProGPT-5.6 SolGPT-5.6 TerraGPT-5.6 Luna
Ceiling
5× Plus
Reachability
—
Pricing notes
Codex limits (Pro 5×, messages): GPT-6 Luna 1,750–14,000 = highest · GPT-6 Sol 70–700 · GPT-5.6 Luna 1,250–10,000 · GPT-5.6 Sol 50–500 · Terra 125–1,000 · Astra 25–225 · weekly limits may also apply
Basic + premium models, covering the mainstream flagships. Non-production use (personal dev / prototyping — PAYG is the production tier). Flow ≈ $0.03283 (floats live); per-model Flow burn and per-tier model lists are not published in text. Provider "Worth" claim: $180/mo equivalent API value (≈1.8× the fee), from its published equivalent-pay-as-you-go table.
GPT-6 AstraGPT-6 SolGPT-6 LunaGPT-5.6 Sol ProGPT-5.6 SolGPT-5.6 TerraGPT-5.6 Luna
Ceiling
20× Plus
Reachability
—
Pricing notes
Codex limits (Pro 20×, messages): GPT-6 Luna 7,000–56,000 = highest · GPT-6 Sol 300–3,000 · GPT-5.6 Luna 5,000–40,000 · GPT-5.6 Sol 200–2,000 · Terra 500–4,000 · Astra 100–900 · weekly limits may also apply
All models (Max and Ultra currently cover nearly identical lists — the difference is quota, not model selection). Non-production use (personal dev / prototyping — PAYG is the production tier). Flow ≈ $0.03283 (floats live); per-model Flow burn and per-tier model lists are not published in text. Provider "Worth" claim: $480/mo equivalent API value (≈2.4× the fee), from its published equivalent-pay-as-you-go table.
Cost rule
n/a
Standard ComputePro$269/mo compute budget · no 5h/week windows · highest-priority scheduling
Grok Bot access is included; linking SuperGrok Heavy grants the highest linked weekly Grok Bot usage on a Cursor account (cursor.com/help/grok-bot/supergrok) — a usage grant, not a Cursor plan, and no extra usage on top of one.
Cost rule
n/a
No plans match your filters.
Some Subscribe buttons carry referral or affiliate codes — using them may give you a signup bonus or support this project at no extra cost to you.
Markers: ~ proxy estimate, ≈ coarse estimate, and ≥ cache floor. A floor means the plan publishes no cached-read rate, so cached tokens bill at the input rate and the real figure is this or higher. Borrowed cap means the budget is a published API cap at list price, so Req / $1 is a value comparison rather than a measured allowance.
Best-priced plans
Best-priced plans for MiMo V2.6 Flash
The top 5 plans by computed Req/$1 that include MiMo V2.6 Flash, plus the lowest-cost plan. Computed at build from the same cost rules the calculator renders · ~ = proxy estimate
“n/a” = the provider publishes no price structure for MiMo V2.6 Flash — the calculator renders the same honest n/a. ◆ best value = highest computed Req/$1.
What real work costs
What MiMo V2.6 Flash costs on real coding tasks
Illustrative cost of common coding tasks on the best-priced plans above, scaled from each plan's live per-request price to a task-typical token volume. Task token assumptions are shared ~proxy estimates — real tasks vary, and they are never per-provider shapes.
Task
CommandCode Go
CommandCode GOAT
OpenCode Go
CommandCode Pro
LLM Gateway DevPass Max
Quick lookup / one-liner8,000 in · 1,000 out
~1¢
~1¢
~1¢
~1¢
~1¢
Review a 500-line PR60,000 in · 4,000 out
~1¢
~1¢
~1¢
~1¢
~1¢
Fix a bug (agent loop)180,000 in · 12,000 out
~1¢
~1¢
~1¢
~1¢
~1¢
Refactor a module320,000 in · 20,000 out
~1¢
~1¢
~1¢
~1¢
~1¢
Full-repo agent run900,000 in · 45,000 out
~1¢
~1¢
~1¢
~1¢
~1¢
Per-task cost = the plan's effective 72,625-token request price scaled to each task's token volume at the same shape mix. ~ = proxy estimate — shared task-token assumptions, not a quote. “n/a” = MiMo V2.6 Flash unpriced on that plan.
Common questions
How many requests per $1 do I get on MiMo V2.6 Flash?
The best plan pricing MiMo V2.6 Flash delivers 15,038 Req/$1 at the model's typical request shape — on OpenCode Go. Figures are exact confidence: they are computed from published rates and usage caps, proxy estimates are marked with a ~, and borrowed-cap plans use a published API cap at list price, so their Req/$1 is a value comparison rather than a measured allowance.
What's the cheapest plan with MiMo V2.6 Flash?
CommandCode Go at $1/mo. The cheapest plan is not always the best value — check Req/$1 above.
Which plans include MiMo V2.6 Flash?
19 tracked plans across 6 providers price it: CommandCode (Go, GOAT, Max 10×, Max 20×, Pro and Provider), Xiaomi MiMo (Lite, Standard, Pro and Max), LLM Gateway DevPass (Lite, Pro and Max), Nous Portal (Plus, Super and Ultra), Ozore (Basic and Pro) and OpenCode Go. The calculator compares every tracked monthly plan that prices it.
Where does MiMo V2.6 Flash rank on the Agent Arena leaderboard?
No verified row yet: MiMo V2.6 Flash is not on the Agent Arena agent leaderboard as of 2026-09-21. Ranks are scraped from https://arena.ai/leaderboard/agent, never typed by hand, so this page shows n/a instead of a guessed position.
What is MiMo V2.6 Flash?
MiMo V2.6 Flash (API mimo-v2.6-flash) is Xiaomi's low-cost reasoning tier in the MiMo V2.6 series (family: xiaomi, kind: text), released 2026-09-22. Xiaomi describes it as full-modality and the best balance for high-frequency calls and large-scale tasks in professional workflows.
How does MiMo V2.6 Flash compare to MiMo V2.6 Pro?
MiMo V2.6 Pro is Xiaomi’s trillion-parameter flagship for complex projects and long-horizon tasks, while MiMo V2.6 Flash is the cheaper full-modality tier for high-frequency and large-scale work; MiMo V2.6 Pro UltraSpeed serves the same Pro model up to 20x faster.
Can I use MiMo V2.6 Flash with any agent (BYOK)?
Yes — the Xiaomi MiMo Token Plan exposes the V2.6 series through a standard endpoint (Xiaomi’s own docs list OpenCode, OpenClaw, Claude Code, Codex and Cline among supported harnesses), and the routers that list Flash serve it over compatible endpoints with no CLI lock.
What is MiMo V2.6 Flash best used for?
High-frequency and large-scale workloads — the volume tier of the MiMo V2.6 line, rather than the flagship pick for long-horizon frontier reasoning.
Newsletter
The weekly brief — plan churn, dated.
A plain weekly digest of every tracked price move, ceiling change and model-list update — pulled straight from the dated changelog ledger, with a one-line “worth switching?” note on each. Free. No spam, unsubscribe anytime.