All plans/Calculator/DeepSeek V4.1 Flash

DeepSeek V4.1 Flash Calculator — requests per $1

Every monthly plan that includes DeepSeek V4.1 Flash, converted to requests per $1 at the model's typical request shape. Same engine as the main calculator — proxy estimates carry a ~ marker; coarse estimates carry ≈.

DeepSeek V4.1 Flash is DeepSeek's newest Flash-generation model for high-volume agentic coding, the successor to DeepSeek V4 Flash in the family's budget-throughput tier.

At DeepSeek V4.1 Flash’s typical shape

3,251req / $1

on OpenCode Go · 33 plans priced · exact

Agent Arena benchmark

#12 (Max) · 4.88% net improvement

Agent Arena leaderboard · fetched 2026-09-21

Per-request shape · tokens

Throttling, off-peak, rollover, and cache writes are not modeled. Where a provider publishes no cached-read price, cached tokens bill at the input rate — those req / $1 figures are floors.

Price–$
36 / 36 plans
Plan detailsProvider and subscription tier. Sorts by provider name.Standard monthly price in USD, before first-month promotions.Estimated cost per request for the selected model and token shape.Requests fitting in the rolling 5-hour budget. '—' means no published 5-hour limit.Requests fitting in the rolling 7-day budget. '—' means no published weekly limit.Requests fitting in 30 days. Weekly budgets are extrapolated; unused credits do not roll over.Requests per $1 of plan price over 30 days. Higher is better.Requests per $1 available in one 5-hour burst. Uncapped plans match Req / $1.SubscribeOpen provider pricing. Referral links may support TokenPlans at no extra cost.
CommandCodeGoCheapest$10 credits · ~15K reqs/mo · $3/5h · $6/7d$1.00$0.00056,50113,00321,673~21,673~6,501Subscribe ↗
CommandCodeGOATNew$70 credits · ~75K reqs/mo · $14/5h · $35/7d$10.00$0.000526,00765,019130,039~13,004~2,601Subscribe ↗
CommandCodePro$80 credits · ~100K reqs/mo · $16/5h · $40/7d$20.00$0.000530,34275,856151,712~7,586~1,517Subscribe ↗
LLM Gateway DevPassMax$537/mo usage at provider rates · weekly frontier fair-use 18%$179.00$0.0005——1,163,849~6,502~6,502Subscribe ↗
LLM Gateway DevPassPro$237/mo usage at provider rates · weekly frontier fair-use 15%$79.00$0.0005——513,654~6,502~6,502Subscribe ↗
LLM Gateway DevPassLite$87/mo usage at provider rates · weekly frontier fair-use 12%$29.00$0.0005——188,556~6,502~6,502Subscribe ↗
CommandCodeMax 20×$300 credits · ~370K reqs/mo · $90/5h · $180/7d$200.00$0.0005195,058390,117650,195~3,251~975Subscribe ↗
CommandCodeMax 10×$150 credits · ~230K reqs/mo · $45/5h · $90/7d$100.00$0.000597,529195,058325,097~3,251~975Subscribe ↗
OpenCodeGo$60/mo usage · $30/wk · $12/5h$10.00$0.00056,50116,25432,5093,251650Subscribe ↗
Standard ComputePro$269/mo compute budget · no 5h/week windows · highest-priority scheduling$249.00$0.0005——583,008~2,341~2,341Subscribe ↗
Standard ComputeStandard$95/mo compute budget · no 5h/week windows · priority scheduling$89.00$0.0005——205,895~2,313~2,313Subscribe ↗
Standard ComputeStarter$20/mo compute budget · no 5h/week windows · one agent at a time$19.00$0.0005——43,346~2,281~2,281Subscribe ↗
Standard ComputeEconomy$41/mo compute budget · no 5h/week windows · parallel agents$39.00$0.0005——88,859~2,278~2,278Subscribe ↗
ZenMuxBuilder Ultra800 Flows / 5h rolling window (16× Starter) · weekly limit · 10-15 RPM$200.00$0.0005——433,463≈2,167≈2,167Subscribe ↗
ZenMuxBuilder Max300 Flows / 5h rolling window (6× Starter) · weekly limit · 10-15 RPM$100.00$0.0005——216,731≈2,167≈2,167Subscribe ↗
ZenMuxBuilder Starter50 Flows / 5h rolling window · weekly limit · 10-15 RPM$20.00$0.0005——43,346≈2,167≈2,167Subscribe ↗
CommandCodeProviderPAYG · top-ups roll over, never expire · zero markup$15.00$0.0005——32,509~2,167~2,167Subscribe ↗
QwenCloudPro Plan-15% off$68/mo40,000 Credits / 7 days$80.00$0.00069,11130,372130,165~1,627~114Subscribe ↗
QwenCloudStandard Plan-28% off$18/mo10,000 Credits / 7 days$25.00$0.00082,2777,59332,541~1,302~91Subscribe ↗
QwenCloudEssential Plan-37.5% off$10/mo5,625 Credits / 7 days$16.00$0.00091,1954,27118,304~1,144~75Subscribe ↗
OllamaPro$60/mo usage credits · access to larger pro models$20.00$0.0009——21,6731,0841,084Subscribe ↗
OllamaMax10 concurrent requests · $300/mo usage credits$100.00$0.0009——108,3651,0841,084Subscribe ↗
QwenCloudLite Plan-25% off$6/mo2,500 Credits / 7 days$8.00$0.00105311,8988,134~1,017~66Subscribe ↗
OzoreBasic$20 in credits/mo · spend on any model at discounted rates · cancel anytime$10.00$0.0077——2,611~≥261~≥261Subscribe ↗
OzorePro$70 in credits/mo · + Smart Compression, Analytics Pro, Privacy Pack, 10 extra API keys, Custom Router, Fusion$35.00$0.0077——9,138~≥261~≥261Subscribe ↗
Phoenix GroveUltra~665M tokens/mo (3.3B+ light) · 7 agents$99.00$0.0043——23,083~233~233Subscribe ↗
Phoenix GroveElite~335M tokens/mo (1.8B+ light) · 6 agents$50.00$0.0043——11,628~233~233Subscribe ↗
Phoenix GroveCanopy~1.3B tokens/mo (6.6B+ light) · 8 agents$195.00$0.0043——45,126~231~231Subscribe ↗
Phoenix GrovePro~165M tokens/mo (890M+ light) · 5 agents$25.00$0.0044——5,727~229~229Subscribe ↗
Phoenix GroveBasic~85M tokens/mo (450M+ light) · 3 agents$12.95$0.0044——2,950~228~228Subscribe ↗
Nous PortalUltra$220 credits/mo · $100 rollover$200.00$0.0099——20,105~≥101~≥101Subscribe ↗
Nous PortalSuper$110 credits/mo · $50 rollover$100.00$0.0099——10,052~≥101~≥101Subscribe ↗
Nous PortalPlus$22 credits/mo · $10 rollover$20.00$0.0099——2,010~≥101~≥101Subscribe ↗
camelAIStreamUnlimited tokens · 1 concurrent generation per stream (extra requests queue — queue waits unbounded, waits can be capped/dropped; 1–50 streams) · 260K context guaranteed$5.00n/a——n/an/an/aSubscribe ↗
ClinePass2-5× standard API rate limits$9.99n/a——n/an/an/aSubscribe ↗
SyntheticSubscription500 requests/5h per pack · 1 concurrent request per model · +500 req/5h per $30 pack (max 5 packs)$30.00n/a——n/an/an/aSubscribe ↗

Some Subscribe buttons carry referral or affiliate codes — using them may give you a signup bonus or support this project at no extra cost to you.

Markers: ~ proxy estimate, ≈ coarse estimate, and ≥ cache floor. A floor means the plan publishes no cached-read rate, so cached tokens bill at the input rate and the real figure is this or higher. Borrowed cap means the budget is a published API cap at list price, so Req / $1 is a value comparison rather than a measured allowance.

Best-priced plans

Best-priced plans for DeepSeek V4.1 Flash

The top 5 plans by computed Req/$1 that include DeepSeek V4.1 Flash, plus the lowest-cost plan. Computed at build from the same cost rules the calculator renders · ~ = proxy estimate

Provider / planPriceReq/$1
CommandCode Go$1◆~21,673
CommandCode GOAT$10~13,004
CommandCode Pro$20~7,586
LLM Gateway DevPass Max$179~6,502
LLM Gateway DevPass Pro$79~6,502

“n/a” = the provider publishes no price structure for DeepSeek V4.1 Flash — the calculator renders the same honest n/a. ◆ best value = highest computed Req/$1.

What real work costs

What DeepSeek V4.1 Flash costs on real coding tasks

Illustrative cost of common coding tasks on the best-priced plans above, scaled from each plan's live per-request price to a task-typical token volume. Task token assumptions are shared ~proxy estimates — real tasks vary, and they are never per-provider shapes.

TaskCommandCode GoCommandCode GOATCommandCode ProLLM Gateway DevPass MaxLLM Gateway DevPass Pro
Quick lookup / one-liner8,000 in · 1,000 out~1¢~1¢~1¢~1¢~1¢
Review a 500-line PR60,000 in · 4,000 out~1¢~1¢~1¢~1¢~1¢
Fix a bug (agent loop)180,000 in · 12,000 out~1¢~1¢~1¢~1¢~1¢
Refactor a module320,000 in · 20,000 out~1¢~1¢~1¢~1¢~1¢
Full-repo agent run900,000 in · 45,000 out~1¢~1¢~1¢~1¢~1¢

Per-task cost = the plan's effective 72,020-token request price scaled to each task's token volume at the same shape mix. ~ = proxy estimate — shared task-token assumptions, not a quote. “n/a” = DeepSeek V4.1 Flash unpriced on that plan.

Common questions

How many requests per $1 do I get on DeepSeek V4.1 Flash?

The best plan pricing DeepSeek V4.1 Flash delivers 3,251 Req/$1 at the model's typical request shape — on OpenCode Go. Figures are exact confidence: they are computed from published rates and usage caps, proxy estimates are marked with a ~, and borrowed-cap plans use a published API cap at list price, so their Req/$1 is a value comparison rather than a measured allowance.

What's the cheapest plan with DeepSeek V4.1 Flash?

CommandCode Go at $1/mo. The cheapest plan is not always the best value — check Req/$1 above.

Which plans include DeepSeek V4.1 Flash?

33 tracked plans across 10 providers price it: CommandCode (Go, GOAT, Max 10×, Max 20×, Pro and Provider), Phoenix Grove (Basic, Pro, Elite, Ultra and Canopy), QwenCloud (Lite Plan, Essential Plan, Standard Plan and Pro Plan), Standard Compute (Starter, Economy, Standard and Pro), LLM Gateway DevPass (Lite, Pro and Max), Nous Portal (Plus, Super and Ultra), ZenMux (Builder Starter, Builder Max and Builder Ultra), Ollama (Pro and Max), Ozore (Basic and Pro) and OpenCode Go. 3 more list it without a published rate, so their calculator rows render n/a: Cline Pass, Synthetic Subscription and camelAI Stream. The calculator compares every tracked monthly plan that prices it.

Where does DeepSeek V4.1 Flash rank on the Agent Arena leaderboard?

#12 (Max config) on the Agent Arena agent leaderboard — 4.88% net improvement (board fetched 2026-09-21). Ranks are scraped from https://arena.ai/leaderboard/agent, never typed by hand.

What is DeepSeek V4.1 Flash?

DeepSeek V4.1 Flash is DeepSeek's newest Flash-generation text generation model (family: deepseek, kind: text), built for coding and agentic work at high volume. It is tracked as its own model rather than as a variant of DeepSeek V4 Flash or DeepSeek V4 Pro.

Can I use DeepSeek V4.1 Flash with any agent (BYOK)?

Yes — the providers above expose DeepSeek V4.1 Flash through OpenAI-compatible endpoints with no CLI lock, so it works with Claude Code, Cursor, OpenCode or any agent that supports custom model endpoints.

How does DeepSeek V4.1 Flash compare to DeepSeek V4 Flash and DeepSeek V4 Pro?

All three are tracked and comparable on tokenplans. DeepSeek V4.1 Flash is the newest generation; V4 Flash stays the cheapest per-request budget tier; V4 Pro is the heavier frontier-reasoning sibling.

Does DeepSeek V4.1 Flash use time-of-day pricing?

On the providers we track it does — DeepSeek V4.1 Flash carries off-peak and peak rates, and which rate a provider stores as the headline varies: most store the provider-published off-peak row, while Ollama stores the peak row. Check the provider page for the peak window if your workload runs during it.

What is DeepSeek V4.1 Flash best used for?

DeepSeek V4.1 Flash is best for high-volume coding and agent loops where cost per request matters but frontier-level reasoning is still required — a step up from the budget V4 Flash tier. For the heaviest multi-step reasoning, pair it with a Pro-tier or frontier model.

Newsletter

The weekly brief — plan churn, dated.

A plain weekly digest of every tracked price move, ceiling change and model-list update — pulled straight from the dated changelog ledger, with a one-line “worth switching?” note on each. Free. No spam, unsubscribe anytime.

Free · weekly · unsubscribe anytime