With your own numbers, no account, in about a minute.
Most teams pay OpenAI or Anthropic per token or per seat. Open-weight models such as DeepSeek, Kimi and GLM do much of the same work for a fraction of that price, hosted in Europe. Below you check that against your own usage: what you pay now, what the same tokens cost here, and what a GPU of your own would do.
You need no account and we store nothing. Pick whatever you have at hand: an admin key (we read your tokens of the last 30 days once and discard the key), your tokens or monthly spend per model, or the subscription you have today.
OpenAI: platform.openai.com, Settings (gear), Organization, Admin keys, Create admin key. You need to be an Owner of the organisation. The key starts with sk-admin and is shown once.
Anthropic: console.anthropic.com, Settings, Admin keys, Create Admin API Key. You need to be an Admin of the organisation. The key starts with sk-ant-admin. Regular keys under API keys (sk-ant-api) do not work.
An admin key can do more than read. Delete it on that same page after the check.
The same traffic on open weights through the router, at today's list prices. Other teams are listed below.
Refreshed daily, converted to euro. Every model gets an open-weight counterpart next to it.
One read, report shown, key gone. We keep neither your tokens nor your subscription.
Your app keeps its OpenAI or Anthropic SDK; only the base URL changes. Behind it you choose per application: shared per token, a dedicated GPU with us, or your own hardware.
Priced with today's list prices and our own rates. Click a case and the form above fills itself; you then get the full report, with every model side by side.
From what we see with customers and in our own numbers. No rule is absolute, which is why the check above exists.
Support, classification, extraction, summaries, translation. This is where the gap is biggest: usually 70 to 95 percent less, and you rarely notice a difference in quality. This is where almost everyone starts.
Agents resend the whole conversation every turn, and Anthropic charges little for those cached tokens. The gap is then smaller than you would expect. Open weights win here with a mid-size model such as DeepSeek V4 Pro or GLM 5.2, not necessarily the very largest.
A team plan costs per person, whether that person uses a lot or a little. Per token you only pay for what actually happens, and for most teams that sits far below the seat price. It also saves the discussion about who gets a seat.
Customer records, HR files, medical intake, insurance claims. On a GPU of your own in the EU with a VPN, your data stays on one machine, no third party sees a prompt and nothing is trained on it. With a processor agreement in place you can process personal data there; the legal basis stays yours, as with any processor.
Example: a claims workflow of 300 million tokens a month on DeepSeek V4 Flash costs €92.40 a month through the router with the EU switch, and €8,753 a month on a dedicated 4x RTX PRO 6000 (392 GB), fully isolated. Which one you need is a compliance choice, not a price choice. More on security
If you rely on one specific strength of a closed frontier model, or if your usage is small. Then the difference is a few euros a month and switching costs more time than it saves. You will see that here too.
From billions of tokens a month on one model. Below that, per token through the router is almost always cheaper, because a machine of your own keeps running at night and at the weekend. The calculator on the pricing page shows the break-even point. To the calculator
You change nothing in your current setup. The check only looks.
An OpenAI or Anthropic admin key, your tokens or monthly spend per model, or the subscription you have today. Whatever you have at hand.
Every model you use gets a comparable open-weight model next to it, per token through our EU router and, when that comes within reach, on a GPU of your own.
Per model what it costs now and what it costs here, plus the same tokens on the popular coding models. Print or save as PDF straight away.
That depends on your workload. For support, classification, extraction and summaries we usually see 70 to 95 percent less. For coding agents with a lot of context the gap is smaller, because Anthropic charges little for cached tokens. For teams on seats the gap is often the biggest, because per token you only pay for what actually happens. The check above works it out with your own usage.
We put GPT-5 next to DeepSeek V4 Pro, Claude Sonnet next to GLM 5.2, Claude Opus next to Kimi K3 and the small models (GPT-5 mini, Claude Haiku) next to DeepSeek V4 Flash. That is a suggestion based on what customers pick in practice, not a benchmark. The report lists every popular coding model side by side with your tokens, so you see at once what a step up or down costs.
Whatever you have at hand. With an OpenAI or Anthropic admin key we read your real usage of the last 30 days. If you know your tokens or monthly spend per model, you enter those. If you pay for a ChatGPT or Claude subscription, you pick the plan and the number of users. All three give the same report.
Only your provider's usage report: tokens per model per day. No prompts, no answers, no end-user names. The key is used once, never for a request to a model, and then discarded. Delete it on your provider's side after the check as well; then the loop is closed.
OpenAI: platform.openai.com, Settings, Organization, Admin keys, Create admin key; you need to be an Owner of the organisation, the key starts with sk-admin. Anthropic: console.anthropic.com, Settings, Admin keys, Create Admin API Key; you need to be an Admin, the key starts with sk-ant-admin. A regular API key cannot read usage.
Your side uses the public list price or seat price, converted to euro; batch discounts or negotiated rates are not included. Our side uses exactly the prices you are billed at here, margin included. Cached tokens count at the cache rate on both sides. Where we are more expensive, it says so.
On price only from billions of tokens a month on one model; below that, per token through the router wins, because a machine of your own keeps running at night. On compliance sooner: whoever processes personal data or regulated work and wants no third party in the chain at all picks a dedicated GPU with a VPN. The report shows a dedicated GPU as soon as it comes close to what you pay now.
Yes, with a processor agreement, as with any processor. Through the router with the EU switch, processing stays with parties established in the EU, all listed on our sub-processors page. On a dedicated GPU with a VPN your data stays on one machine and nobody else sees a prompt. In both cases nothing is trained on your data. The legal basis for the processing stays yours.
No. You compare right here, without an account. You only switch once the numbers convince you, and it does not have to happen all at once: the router speaks both the OpenAI and the Anthropic API, so you can switch one application at a time.
It takes a minute and we keep nothing. Switching is a separate decision after that, one you make with the numbers in hand. Not sure about a model? We are happy to look with you.