DeepFrugal Early Preview Find the cheapest way to run any model

Subscription vs pay-as-you-go: how to find your break-even

A subscription only pays off above a usage threshold. Learn how DeepFrugal computes effective subscription prices, including cached tokens, and find your break-even with the interactive calculator.

12 min read DeepFrugal
Subscription vs pay-as-you-go: how to find your break-even

A flat monthly plan feels cheap next to a pay-as-you-go API. It is not always cheaper. The plan wins only above a usage threshold. Below that threshold you pay for capacity you never use.

This guide shows how DeepFrugal turns a plan's list price into an effective price per token, and how to find the usage where a plan beats paying per token.

Why the list price is not the real price

Gateways publish prices in three shapes. DeepFrugal normalises all of them to $ per 1M tokens.

Billing Real price
Pay as you go The listed price.
Subscription Listed ร— (subscription รท quota).
Reseller Listed ร— (1 + service fee). Sales tax is not included.

Many plans grant a usage quota instead of a fixed token count. The quota takes a few shapes:

  • Dollar credit โ€” a budget of dollars spent at the listed rates. A plan that buys a multiple of what you pay (3ร— the subscription, say) is the same kind, just a bigger budget.
  • Token credit โ€” the plan grants credits, and each model consumes a different number of them per token.
  • Usage limited โ€” messages, tasks or a shared pool, with no token or credit amount published.

A dollar-credit quota is a budget, not a token count. It is spent at a rate that depends on the model: an expensive model consumes more of the budget for the same listed dollars. DeepFrugal scales the listed rate by the plan, so a single model's effective rate depends only on the subscription and its quota. A mix spends the budget in proportion to each model's share. For the full picture, see how AI subscription credits work.

Listed prices and quotas change over time. This table shows one example of a plan's listed and effective rates:

Listed vs effective rates โ€” OpenCode Go
Subscription $10.00/month ยท quota $60.00/month for this model ยท effective = listed ร— $10.00 รท $60.00
Input /1MOutput /1MCached read /1M
$0.025$0.15$0.1$0.6$0.001$0.003

Figures are examples. Plans, promos and prices change over time. Check the live table for current values.

The effective rate is a best case. It assumes you consume every included token on the selected model. DeepFrugal states this in every price tooltip.

Don't forget cached tokens

Most providers bill cached input at a lower rate than fresh input, and some charge a separate cache write fee. A cached read is often around 10% of the input price. If your workload reuses a long prompt, cache reads can dominate your bill.

Three token types therefore drive the real cost:

  • Input โ€” fresh prompt tokens.
  • Cached read โ€” prompt tokens served from cache.
  • Cached write โ€” tokens written into the cache, when the provider charges.
  • Output โ€” generated tokens.

The calculator below includes all of them. Add your monthly cached input and cache writes to see how much the subscription really saves.

Find your break-even

Pick a plan, then add the models you actually use. Each row has its own monthly input, cached read, cached write and output; use Add model to build a mix. The default is the Medium preset. Pick Light, Medium or Heavy ยท agentic to load a representative mix, then edit it; each chip names the models it will load for the selected plan.

Plan break-even โ€” OpenCode Go ยท Experimental
Cheaper
Subscriptionsave $24.19/month
You pay $10.00/month instead of $34.19 of usage.
Usage: 663M tokens ยท $50.35 at list rates ยท $34.19 cheapest pay-as-you-go.
Subscriptioncheaper$10.00/month
Quota used84%
Covered by the quota$50.35
Above the quota$0.00
Monthly cost$10.00
Cheapest pay-as-you-go route

Real rates include the service fee; sales tax is not included.

GLM 5.3 Flash via Ozore API

DeepSeek V4.1 Flash via DeepSeek API

Monthly cost$34.19
Break-even: $10.00 of usage (โ‰ˆ 132M tokens at this mix)
At your mix: subscription $10.00/mo ยท cheapest metered $34.19/mo ยท at list rates $50.35/mo โ€” subscription is cheapest by $24.19/mo.
Plan break-even details
$0.013$0.015$0.052$0.076500M1B663M132M194M2.05B790MTotal tokens per month (M tokens/mo, log scale)Average cost per 1M tokens overflow billed at the listed rate your mix
  • Average cost per token
  • Listed rate
  • Cheapest pay-as-you-go route
  • Quota limit
  • Cheaper than the cheapest
  • Quota usage break-even
  • Plan break-even / loses
  • Average at your mix
Hover the curve to read the values.
โ–พ How the graph is calculated

Inside the quota the average is the monthly commitment divided by the tokens used, so it falls as the mix scales up and reaches the mix's effective rate at the quota limit. Above the limit the budget is spent and the extra tokens are billed at the overflow base chosen in Options โ€” the model's listed rate (default), the cheapest pay-as-you-go rate found for it, or the cheaper of the two โ€” so the average climbs towards that blended rate.

The quota is one shared budget: each model's published quota is that same budget restated at its rates, and the budget is allocated in the model order shown, so a different order or blend moves the curve. A token type a model does not price still counts in the volume but adds no cost. The token axis is logarithmic, so a wide range of usage fits on one chart.

The pay-as-you-go line is the cheapest real rate per model (fee included; sales tax excluded) โ€” an exact variant match when a pay-as-you-go gateway publishes that variant, otherwise the model's nearest published row โ€” so it can combine more than one gateway. When that rate equals the plan's own listed rate the chart shows the listed line only.

The markers: the quota usage break-even is where the average meets the plan's list rate, so below it you pay more per token than list; the plan break-even is where the average starts to beat the cheaper of the two references; loses to the cheapest is where it stops doing so (it exists only when the overflow base is dearer than that reference); the your mix point is your stated usage, labelled with its monthly token volume, where the horizontal guide marks the mix's average cost per 1M, and the shaded band is where the subscription wins. With Promo prices off, every figure uses the pre-promo list rates and quotas, and the commitment reverts to the list subscription.

โ–พ How the plan quota is consumed

A subscription grants one monthly budget. The budget is allocated to the models in the order listed here: each model spends its usage divided by its quota, capped by what is left, so the total never exceeds 100%. Reorder the models to approximate your own pattern. Usage above a model's quota, or beyond the budget, is billed at the listed rate.

ModelUsageMonthly quota% of QuotaPay as you go
GLM 5.3 Flash$29.10$60.0048.5%โ€”
DeepSeek V4.1 Flash (Off-Peak)$15.71$60.0026.2%โ€”
DeepSeek V4.1 Flash (Peak)$5.54$60.009.2%โ€”
Total$50.35โ€”84%$0.00

Budget used = 0.839 (84%) ยท Covered = $50.35 ยท Pay as you go = $0.00

โ–พ Effective rates (listed vs effective)
ModelInput /1MOutput /1MCached read /1M
GLM 5.3 Flash$0.03$0.15$0.099$0.5$0.006$0.03
DeepSeek V4.1 Flash (Off-Peak)$0.03$0.15$0.119$0.6$0.001$0.003
DeepSeek V4.1 Flash (Peak)$0.06$0.3$0.238$1.2$0.001$0.006

Effective = listed ร— 0.199 at this usage ยท 84% of the plan budget used

Promo active

These figures use promotional rates or quotas. They change when the promo ends โ€” confirm current pricing before you decide.

  • DeepSeek V4.1 Flash (Off-Peak) โ€” 4ร— usage promo ($15 โ†’ $60 quota) ยท ends 2026-09-20
  • DeepSeek V4.1 Flash (Peak) โ€” 4ร— usage promo ($15 โ†’ $60 quota) ยท ends 2026-09-20

Experimental. Confirm every figure against the provider's own pricing before you rely on it. DeepFrugal is not responsible for calculation errors.

The same calculator is available as a standalone page at /tools/plan-break-even/. It keeps your usage in the URL, so you can bookmark a calculation or share it.

The calculator values your usage at the listed rates and compares it with the plan. Below the subscription price, pay-as-you-go is cheaper. Between the subscription price and the quota, you pay the flat subscription and the rest is saving.

The plan grants one monthly budget for the whole mix. We cannot know the order in which you draw from each model, so the calculator assumes the list order; use the arrows to approximate your own pattern. Usage above a model's own quota, or beyond the budget, is not covered: choose what the overflow is billed at โ€” the model's listed rate (the default), the cheapest pay-as-you-go rate found for it, or the cheaper of the two.

The calculator reads the same data as the main table. It recomputes on every change.

Open How the plan quota is consumed to see each model's share of the plan quota and its pay-as-you-go excess, and Effective rates (listed vs effective) for the listed-versus-effective table. Both panels start collapsed to keep the calculator compact.

Compare providers for one model

Model prices differ by provider. The same model can cost several times more on one endpoint than another. Peak, off-peak and flex tiers are separate rows.

Cheapest providers
PlanPricing /1MMonthly quota
InputOutputCached read
Ozore APIOzore$0.11$0.42โ€”โ€”
Ozore BasicOzore$0.11$0.42โ€”$20 of usage credits
Ozore ProOzore$0.11$0.42โ€”$70 of usage credits
OpenRouter APIOpenRouter$0.13$0.52$0.003โ€”
Portal PlusNous Portal$0.13$0.52$0.003$22 of usage credits ยท $10 rollover cap
Portal PlusNous Portal$0.13$0.52$0.003$22 of usage credits ยท $10 rollover cap
Portal SuperNous Portal$0.13$0.52$0.003$110 of usage credits ยท $50 rollover cap
Portal SuperNous Portal$0.13$0.52$0.003$110 of usage credits ยท $50 rollover cap
Portal UltraNous Portal$0.13$0.52$0.003$220 of usage credits ยท $100 rollover cap
Portal UltraNous Portal$0.13$0.52$0.003$220 of usage credits ยท $100 rollover cap

The table ranks providers by their listed rate โ€” the price you pay per token on pay-as-you-go, before any subscription scaling or fee. Each row links to the gateway. Turn on Effective prices to see the rate with the fee applied, or scaled down by a plan's quota; the ranking updates with it. Sales tax is never included.

Browse a plan

Subscription plans bundle many models. This table lists a plan's models with their effective prices, cached reads and writes, and the monthly quota each model draws from. Use the selector to switch plan, and click a column header to sort.

Effective prices โ€” OpenCode Go
ModelVariantInput /1MOutput /1MCached read /1MCached write /1MMonthly quota
Muse Spark 1.3 ContributorMuse Spark 1.3 Contributor$0.017$0.033$0โ€”$60
Muse Spark 1.2 ContributorMuse Spark 1.2 Contributor$0.017$0.033$0โ€”$60
MiMo-V2.5MiMo-V2.5$0.023$0.047$0โ€”$60
Hy3Hy3$0.023$0.097$0.006โ€”$60
GLM 5.3 FlashGLM 5.3 Flash$0.025$0.083$0.005โ€”$60
DeepSeek V4.1 FlashDeepSeek V4.1 Flash (Off-Peak)$0.025$0.1$0.001โ€”$60$15
LongCat-2.0LongCat-2.0$0.05$0.2$0.001โ€”$60
MiniMax M3MiniMax M3$0.05$0.2$0.01โ€”$60
MiniMax M2.7MiniMax M2.7$0.05$0.2$0.01$0.063$60
Qwen3.8 FlashQwen3.8 Flash$0.05$0.157$0.005$0.067$30
DeepSeek V4.1 FlashDeepSeek V4.1 Flash (Peak)$0.05$0.2$0.001โ€”$60$15
DeepSeek V4 FlashDeepSeek V4 Flash (Off-Peak)$0.05$0.2$0.001โ€”$30
Qwen3.7 PlusQwen3.7 Plus (โ‰ค 256K tokens)$0.067$0.267$0.007$0.083$60
Qwen3.6 PlusQwen3.6 Plus (โ‰ค 256K tokens)$0.083$0.5$0.008$0.104$60
DeepSeek V4 FlashDeepSeek V4 Flash (Peak)$0.1$0.4$0.002โ€”$30
DeepSeek V4 Flash VisionDeepSeek V4 Flash Vision Exp (Off-Peak)$0.1$0.4$0.002โ€”$15
GPT-5.6 LunaGPT-5.6 Luna (โ‰ค 272K tokens)$0.133$0.8$0.013$0.167$15
Kimi K2.7 CodeKimi K2.7 Code$0.158$0.667$0.032โ€”$60
Kimi K2.6Kimi K2.6$0.158$0.667$0.027โ€”$60
Qwen3.7 PlusQwen3.7 Plus (> 256K tokens)$0.2$0.8$0.02$0.25$60
DeepSeek V4 Flash VisionDeepSeek V4 Flash Vision Exp (Peak)$0.2$0.8$0.004โ€”$15
GLM-5.2GLM-5.2$0.233$0.733$0.043โ€”$60
GLM-5.1GLM-5.1$0.233$0.733$0.043โ€”$60
GPT-5.6 LunaGPT-5.6 Luna (> 272K tokens)$0.267$1.2$0.027$0.333$15
Hy4Hy4 Preview$0.278$0.834$0.014โ€”$30
MiMo-V2.5-ProMiMo-V2.5-Pro$0.29$0.58$0.002โ€”$15
Qwen3.6 PlusQwen3.6 Plus (> 256K tokens)$0.333$1$0.033$0.417$60
DeepSeek V4 ProDeepSeek V4 Pro (Off-Peak)$0.44$1.32$0.015โ€”$15
Qwen3.7 MaxQwen3.7 Max$0.833$2.5$0.167$1.042$30
DeepSeek V4 ProDeepSeek V4 Pro (Peak)$0.88$2.64$0.029โ€”$15
GLM 5.3GLM 5.3$0.933$2.933$0.173โ€”$15
Qwen3.8 MaxQwen3.8 Max$1.333$4$0.167$1.667$15
Grok 4.6Grok 4.6 (โ‰ค 200K tokens)$1.333$4$0.333โ€”$15
Kimi K3Kimi K3$2$10$0.2โ€”$15
Grok 4.6Grok 4.6 (> 200K tokens)$2.667$8$0.667โ€”$15

Read the live table

Every widget links back to the live comparison table. There you can filter by provider, context, latency, prompt logging and more.

A few rules of thumb:

  • Low usage: pay as you go. You pay only for what you send.
  • High, steady usage: a subscription. The effective rate falls as you use the quota.
  • Cache-heavy usage: check the cached read rate. It can change the ranking.
  • Bursty usage: check the daily and weekly quotas. Many plans throttle bursts well below the monthly quota.
  • Resellers: add the service fee before you compare. Sales tax is not included โ€” confirm it with the provider.

Summary

The cheapest plan is the one that matches your usage. DeepFrugal shows both numbers side by side, so you can decide with data instead of marketing copy.

Frequently asked questions

When does a subscription beat pay-as-you-go?

Above a usage threshold, once the flat commitment is spread over enough tokens. Below it, metered billing is cheaper.

What is an effective price per token?

The listed rate scaled by the plan: listed multiplied by the subscription and divided by the quota. It assumes the whole quota is used.

How do cached tokens change the result?

Cached reads are usually cheaper and cache writes may cost extra. The calculator includes input, cached read, cached write and output.

What if I use more than the quota?

The excess is billed separately, at the plan's overflow base โ€” usually the model's listed rate. The plan covers the quota; you pay for the rest.

Do weekly and daily limits matter?

Yes for bursty usage. They cap how fast you can spend the budget and can throttle you before the month ends.