Usage & Rate Card

Every plan includes a monthly usage level — 1x on Starter, 5x on Team, 10x on Business. This page shows what a level covers and what each model, page and minute costs.

How usage works

  • One allowance per organisation. The plan's level is a monthly allowance that every member, API key, agent and surface installation in your organisation draws from.
  • Charged when it completes. A model call, a converted page or a transcribed minute is charged once it has run, even when the answer is not what you wanted.
  • The model you pick matters most. The standard built-in model is a small, fast one; frontier models use the allowance visibly faster. The worked examples show by how much.
  • Caching can lower costs. Cached input tokens are charged at a lower rate, as the rate card shows. How much you save depends on your workload, so no figure is promised.
  • Bring your own keys. On Team, Business and Enterprise, a request routed to your own provider is billed by that provider, and its model charge does not count against your plan's usage. Other charges, such as pages and transcription, still do.
  • It resets every month. Usage is granted monthly, on annual plans too, and unused usage does not carry over. The console shows how much of this month's usage is used and when it resets.
  • Extra usage is opt-in. When the allowance is used up, requests are refused with 402 unless an admin has turned on extra usage or bought a top-up.

Worked examples

What one run of each workload costs at the rate card prices, rounded to four decimals, and how much of each usage level it takes: the share of the level’s monthly usage one run uses and, in brackets, about how many runs the level covers in a month if nothing else draws on it. The examples assume no cached input; cached tokens cost less.

  • Chat reply: one call, 8,000 input and 500 output tokens.
  • Agent run: 10 calls, each 15,000 input and 1,000 output tokens.
  • Long agent run: 50 calls, each 100,000 input and 2,000 output tokens.
WorkloadModelCost per run1x (Starter)5x (Team)10x (Business)
Chat replygpt-5.1-mini€0.01130.025% (about 4,000)0.005% (about 20,000)0.0025% (about 40,000)
Chat replygpt-5.1€0.05630.13% (about 800)0.025% (about 4,000)0.013% (about 8,000)
Chat replygpt-5.4€0.11340.25% (about 396)0.05% (about 1,983)0.025% (about 3,966)
Chat replygpt-5.6-sol€0.17330.39% (about 259)0.077% (about 1,298)0.038% (about 2,597)
Agent rungpt-5.1-mini€0.21560.48% (about 208)0.096% (about 1,043)0.048% (about 2,086)
Agent rungpt-5.1€1.07812.4% (about 41)0.48% (about 208)0.24% (about 417)
Agent rungpt-5.4€2.16564.8% (about 20)0.96% (about 103)0.48% (about 207)
Agent rungpt-5.6-sol€3.307.3% (about 13)1.5% (about 68)0.73% (about 136)
Long agent rungpt-5.1-mini€5.437512% (about 8)2.4% (about 41)1.2% (about 82)
Long agent rungpt-5.1€27.187560% (about 1)12% (about 8)6% (about 16)
Long agent rungpt-5.4€57.75128% (fewer than 1)26% (about 3)13% (about 7)
Long agent rungpt-5.6-sol€90.75202% (fewer than 1)40% (about 2)20% (about 4)

Rate card for developers

The console's model list and GET /v1/models are authoritative for what your organisation can call; this table prices the built-in models. Each call is charged on its own, for the tokens it used. Extra usage is billed at these prices.

Euros per 1M tokens, net of VAT. These prices are in effect since 2026-10-01.

ModelInputCached inputCache writeOutput
gpt-4.1€7.50€1.875€7.50€30.00
gpt-4.1-mini€1.50€0.375€1.50€6.00
gpt-4.1-nano€0.375€0.09375€0.375€1.50
gpt-5.1€4.6875€0.46875€4.6875€37.50
gpt-5.1-mini€0.9375€0.09375€0.9375€7.50
gpt-5.4€10.3125€1.03125€10.3125€61.875
gpt-5.6-sol€16.50€1.65€20.625€82.50
gpt-5.6-terra€8.25€0.825€10.3125€49.50
gpt-5.6-luna€0.825€0.0825€1.03125€4.95
gpt-6.1-sol€9.00€0.45€11.25€45.00
gpt-6-luna€0.45€0.045€0.5625€2.25

A call whose input exceeds 272,000 tokens is charged at the long-context rate for the whole call:

ModelInputCached inputCache writeOutput
gpt-5.4€20.625€2.0625€20.625€92.8125
gpt-5.6-sol€33.00€3.30€41.25€123.75
gpt-5.6-terra€16.50€1.65€20.625€74.25
gpt-5.6-luna€1.65€0.165€2.0625€7.425
gpt-6.1-sol€18.00€0.90€22.50€67.50
gpt-6-luna€0.90€0.09€1.125€3.375

Embeddings, euros per 1M input tokens:

ModelInput
text-embedding-3-small€0.0825
text-embedding-3-large€0.53625

Everything else:

ResourceUnitPrice
Document conversionper page€0.015
Audio transcriptionper minute€0.0225
Dictationper hour€3.75
Web searchper search€0.01875
Web fetchper page fetched€0.0015

Extra usage and top-ups

  • Extra usage is off by default. An admin turns it on and sets a monthly cap in euros, up to €100 a month. It is charged at the euro prices on the rate card and invoiced monthly in arrears. An organisation with an overdue invoice gets no new extra usage until that invoice is paid.
  • Top-ups are prepaid through checkout. Each top-up stays valid for 12 months after purchase and is used before extra usage.
  • Prices, caps and top-ups are net of VAT.
  • Enterprise contracts set their own usage, cap and terms.

Was this page helpful?