Model garden Router · warm Dedicated · available

gpt oss 120b

Instantly via the EU router or as a dedicated GPU deployment. Data stays in Europe.

Welcome to the gpt-oss series, OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases.

openai/gpt-oss-120b
text->text · openai · sovereign EU
Runs on EU infrastructure operated by European companies; US marketplace capacity is never part of this chain. Full chain
117B
Parameters
131K
Context window
160GB
Minimum VRAM
POST /api/v1/chat/completions 200 OK

Specifications

Parameters 117B
Context window 131,072 tokens
Minimum VRAM 160 GB
Architecture GptOssForCausalLM (vLLM)
License apache-2.0
Modality text->text
Released August 2025
Publisher openai ↗

Pricing

Shared router · per token
€0.16
Input (per 1M tokens)
€0.63
Output (per 1M tokens)
Dedicated GPU · per hour
from €5,07 per hour
Your own vLLM instance on European cloud (160 GB VRAM), billed hourly.

Shared EU router, pay-per-token, scale-to-zero. Dedicated GPU deployments are billed hourly, see pricing.

Call it now

Drop-in replacement for OpenAI: change only the base URL and API key. The Anthropic format (/v1/messages) is supported too.

curl https://hostyourai.com/api/v1/chat/completions \
  -H "Authorization: Bearer hyai-..." \
  -H "Content-Type: application/json" \
  -d '{
    "model": "openai/gpt-oss-120b",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Frequently asked questions

Can I run gpt oss 120b in the EU?

Yes. HostYourAI runs gpt oss 120b on GPUs in European datacenters via vLLM. Prompts and outputs never leave the EU and there is no US cloud provider in the chain.

Is hosting gpt oss 120b GDPR-compliant?

Yes. All processing happens inside the EU, a Data Processing Agreement (DPA) is available and the subprocessor list is public. Open-source weights also mean: no training on your data.

How much does gpt oss 120b cost?

Via the shared EU router you pay €0.16 per million input tokens and €0.63 per million output tokens, with no fixed costs. For high volume or isolation you can also run gpt oss 120b as a dedicated hourly GPU instance.

Is the API OpenAI-compatible?

Yes. You use the standard OpenAI SDKs with a custom base URL (https://hostyourai.com/api/v1). The Anthropic Messages API is supported as a drop-in as well.

More models from OpenAI

gpt oss safeguard 20b

gpt-oss-safeguard-120b and gpt-oss-safeguard-20b are safety reasoning models built-upon gpt-oss. With these models, you can classify text content based on safety policies that you provide and perform a suite of foundational safety tasks. These models are intended for safety use cases. For other applications, we recommend using gpt-oss models.

22B 131K context View model →
gpt oss safeguard 120b

gpt-oss-safeguard-120b and gpt-oss-safeguard-20b are safety reasoning models built-upon gpt-oss. With these models, you can classify text content based on safety policies that you provide and perform a suite of foundational safety tasks. These models are intended for safety use cases. For other applications, we recommend using gpt-oss models.

120B 131K context View model →
gpt oss 20b

Welcome to the gpt-oss series, OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases.

21B 131K context View model →
whisper large v3 turbo

Whisper is a state-of-the-art model for automatic speech recognition (ASR) and speech translation, proposed in the paper Robust Speech Recognition via Large-Scale Weak Supervision by Alec Radford et al. from OpenAI. Trained on 5M hours of labeled data, Whisper demonstrates a strong ability to generalise to many datasets and domains in a zero-shot setting.

0.8B View model →
whisper large v3

Whisper is a state-of-the-art model for automatic speech recognition (ASR) and speech translation, proposed in the paper Robust Speech Recognition via Large-Scale Weak Supervision by Alec Radford et al. from OpenAI. Trained on 5M hours of labeled data, Whisper demonstrates a strong ability to generalise to many datasets and domains in a zero-shot setting.

1.5B View model →
whisper large v2

Whisper is a pre-trained model for automatic speech recognition (ASR) and speech translation. Trained on 680k hours of labelled data, Whisper models demonstrate a strong ability to generalise to many datasets and domains without the need for fine-tuning.

1.5B View model →

Try gpt oss 120b for free

Creating an account takes a minute. Test gpt oss 120b straight away in the playground.

Start for free