Usage & Rate Card
Every plan includes a monthly usage level — 1x on Starter, 5x on Team, 10x on Business. This page shows what a level covers and what each model, page and minute costs.
How usage works
- One allowance per organisation. The plan's level is a monthly allowance that every member, API key, agent and surface installation in your organisation draws from.
- Charged when it completes. A model call, a converted page or a transcribed minute is charged once it has run, even when the answer is not what you wanted.
- The model you pick matters most. The standard built-in model is a small, fast one; frontier models use the allowance visibly faster. The worked examples show by how much.
- Caching can lower costs. Cached input tokens are charged at a lower rate, as the rate card shows. How much you save depends on your workload, so no figure is promised.
- Bring your own keys. On Team, Business and Enterprise, a request routed to your own provider is billed by that provider, and its model charge does not count against your plan's usage. Other charges, such as pages and transcription, still do.
- It resets every month. Usage is granted monthly, on annual plans too, and unused usage does not carry over. The console shows how much of this month's usage is used and when it resets.
- Extra usage is opt-in. When the allowance is used up, requests are refused with
402unless an admin has turned on extra usage or bought a top-up.
Worked examples
What one run of each workload costs at the rate card prices, rounded to four decimals, and how much of each usage level it takes: the share of the level’s monthly usage one run uses and, in brackets, about how many runs the level covers in a month if nothing else draws on it. The examples assume no cached input; cached tokens cost less.
- Chat reply: one call, 8,000 input and 500 output tokens.
- Agent run: 10 calls, each 15,000 input and 1,000 output tokens.
- Long agent run: 50 calls, each 100,000 input and 2,000 output tokens.
| Workload | Model | Cost per run | 1x (Starter) | 5x (Team) | 10x (Business) |
|---|---|---|---|---|---|
| Chat reply | gpt-5.1-mini | €0.0113 | 0.025% (about 4,000) | 0.005% (about 20,000) | 0.0025% (about 40,000) |
| Chat reply | gpt-5.1 | €0.0563 | 0.13% (about 800) | 0.025% (about 4,000) | 0.013% (about 8,000) |
| Chat reply | gpt-5.4 | €0.1134 | 0.25% (about 396) | 0.05% (about 1,983) | 0.025% (about 3,966) |
| Chat reply | gpt-5.6-sol | €0.1733 | 0.39% (about 259) | 0.077% (about 1,298) | 0.038% (about 2,597) |
| Agent run | gpt-5.1-mini | €0.2156 | 0.48% (about 208) | 0.096% (about 1,043) | 0.048% (about 2,086) |
| Agent run | gpt-5.1 | €1.0781 | 2.4% (about 41) | 0.48% (about 208) | 0.24% (about 417) |
| Agent run | gpt-5.4 | €2.1656 | 4.8% (about 20) | 0.96% (about 103) | 0.48% (about 207) |
| Agent run | gpt-5.6-sol | €3.30 | 7.3% (about 13) | 1.5% (about 68) | 0.73% (about 136) |
| Long agent run | gpt-5.1-mini | €5.4375 | 12% (about 8) | 2.4% (about 41) | 1.2% (about 82) |
| Long agent run | gpt-5.1 | €27.1875 | 60% (about 1) | 12% (about 8) | 6% (about 16) |
| Long agent run | gpt-5.4 | €57.75 | 128% (fewer than 1) | 26% (about 3) | 13% (about 7) |
| Long agent run | gpt-5.6-sol | €90.75 | 202% (fewer than 1) | 40% (about 2) | 20% (about 4) |
Rate card for developers
The console's model list and GET /v1/models are authoritative for what your organisation can call; this table prices the built-in models. Each call is charged on its own, for the tokens it used. Extra usage is billed at these prices.
Euros per 1M tokens, net of VAT. These prices are in effect since 2026-10-01.
| Model | Input | Cached input | Cache write | Output |
|---|---|---|---|---|
gpt-4.1 | €7.50 | €1.875 | €7.50 | €30.00 |
gpt-4.1-mini | €1.50 | €0.375 | €1.50 | €6.00 |
gpt-4.1-nano | €0.375 | €0.09375 | €0.375 | €1.50 |
gpt-5.1 | €4.6875 | €0.46875 | €4.6875 | €37.50 |
gpt-5.1-mini | €0.9375 | €0.09375 | €0.9375 | €7.50 |
gpt-5.4 | €10.3125 | €1.03125 | €10.3125 | €61.875 |
gpt-5.6-sol | €16.50 | €1.65 | €20.625 | €82.50 |
gpt-5.6-terra | €8.25 | €0.825 | €10.3125 | €49.50 |
gpt-5.6-luna | €0.825 | €0.0825 | €1.03125 | €4.95 |
gpt-6.1-sol | €9.00 | €0.45 | €11.25 | €45.00 |
gpt-6-luna | €0.45 | €0.045 | €0.5625 | €2.25 |
A call whose input exceeds 272,000 tokens is charged at the long-context rate for the whole call:
| Model | Input | Cached input | Cache write | Output |
|---|---|---|---|---|
gpt-5.4 | €20.625 | €2.0625 | €20.625 | €92.8125 |
gpt-5.6-sol | €33.00 | €3.30 | €41.25 | €123.75 |
gpt-5.6-terra | €16.50 | €1.65 | €20.625 | €74.25 |
gpt-5.6-luna | €1.65 | €0.165 | €2.0625 | €7.425 |
gpt-6.1-sol | €18.00 | €0.90 | €22.50 | €67.50 |
gpt-6-luna | €0.90 | €0.09 | €1.125 | €3.375 |
Embeddings, euros per 1M input tokens:
| Model | Input |
|---|---|
text-embedding-3-small | €0.0825 |
text-embedding-3-large | €0.53625 |
Everything else:
| Resource | Unit | Price |
|---|---|---|
| Document conversion | per page | €0.015 |
| Audio transcription | per minute | €0.0225 |
| Dictation | per hour | €3.75 |
| Web search | per search | €0.01875 |
| Web fetch | per page fetched | €0.0015 |
Extra usage and top-ups
- Extra usage is off by default. An admin turns it on and sets a monthly cap in euros, up to €100 a month. It is charged at the euro prices on the rate card and invoiced monthly in arrears. An organisation with an overdue invoice gets no new extra usage until that invoice is paid.
- Top-ups are prepaid through checkout. Each top-up stays valid for 12 months after purchase and is used before extra usage.
- Prices, caps and top-ups are net of VAT.
- Enterprise contracts set their own usage, cap and terms.