Model garden Router · on request Dedicated · available

whisper small.en

Runs as your own dedicated GPU deployment, ready to launch from the wizard. Data stays in Europe.

Whisper is a pre-trained model for automatic speech recognition (ASR) and speech translation. Trained on 680k hours of labelled data, Whisper models demonstrate a strong ability to generalise to many datasets and domains without the need for fine-tuning.

openai/whisper-small.en On request
audio->text · openai · sovereign EU
Runs on EU infrastructure operated by European companies; US marketplace capacity is never part of this chain. Full chain
0.2B
Parameters
Context window
8GB
Minimum VRAM
POST /api/v1/audio/transcriptions On request

Specifications

Parameters 0.2B
Minimum VRAM 8 GB
Architecture WhisperForConditionalGeneration (vLLM)
License apache-2.0
Modality audio->text
Released September 2022
Publisher openai ↗

Pricing

Shared router · per token
On request
Not available on the shared router. Pricing on request as a dedicated GPU deployment.
Dedicated GPU · per hour
from €1,95 per hour
Your own vLLM instance on European cloud (8 GB VRAM), billed hourly.

Shared EU router, pay-per-token, scale-to-zero. Dedicated GPU deployments are billed hourly, see pricing.

Call it now

Drop-in replacement for OpenAI: change only the base URL and API key. The Anthropic format (/v1/messages) is supported too.

curl https://hostyourai.com/api/v1/audio/transcriptions \
  -H "Authorization: Bearer hyai-..." \
  -F model="openai/whisper-small.en" \
  -F file="@meeting.mp3"

Frequently asked questions

Can I run whisper small.en in the EU?

Yes. HostYourAI runs whisper small.en on GPUs in European datacenters via vLLM. Prompts and outputs never leave the EU and there is no US cloud provider in the chain.

Is hosting whisper small.en GDPR-compliant?

Yes. All processing happens inside the EU, a Data Processing Agreement (DPA) is available and the subprocessor list is public. Open-source weights also mean: no training on your data.

How much does whisper small.en cost?

whisper small.en needs several GPUs at once, so it runs as a dedicated deployment billed per GPU-hour rather than per token. Tell us your volume and we will work it out with you.

Is the API OpenAI-compatible?

Yes. You use the standard OpenAI SDKs with a custom base URL (https://hostyourai.com/api/v1). The Anthropic Messages API is supported as a drop-in as well.

More models from OpenAI

gpt oss safeguard 20b

gpt-oss-safeguard-120b and gpt-oss-safeguard-20b are safety reasoning models built-upon gpt-oss. With these models, you can classify text content based on safety policies that you provide and perform a suite of foundational safety tasks. These models are intended for safety use cases. For other applications, we recommend using gpt-oss models.

22B 131K context View model →
gpt oss safeguard 120b

gpt-oss-safeguard-120b and gpt-oss-safeguard-20b are safety reasoning models built-upon gpt-oss. With these models, you can classify text content based on safety policies that you provide and perform a suite of foundational safety tasks. These models are intended for safety use cases. For other applications, we recommend using gpt-oss models.

120B 131K context View model →
gpt oss 20b

Welcome to the gpt-oss series, OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases.

22B 131K context View model →
gpt oss 120b

Welcome to the gpt-oss series, OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases.

117B 131K context View model →
whisper large v3 turbo

Whisper is a state-of-the-art model for automatic speech recognition (ASR) and speech translation, proposed in the paper Robust Speech Recognition via Large-Scale Weak Supervision by Alec Radford et al. from OpenAI. Trained on 5M hours of labeled data, Whisper demonstrates a strong ability to generalise to many datasets and domains in a zero-shot setting.

0.8B View model →
whisper large v3

Whisper is a state-of-the-art model for automatic speech recognition (ASR) and speech translation, proposed in the paper Robust Speech Recognition via Large-Scale Weak Supervision by Alec Radford et al. from OpenAI. Trained on 5M hours of labeled data, Whisper demonstrates a strong ability to generalise to many datasets and domains in a zero-shot setting.

1.5B View model →

Try whisper small.en for free

Creating an account takes a minute. Your API key works with whisper small.en right away.

Start for free