Model garden Router · op aanvraag Dedicated · beschikbaar

paza whisper large v3 turbo

Dit model draait als eigen dedicated GPU-deployment, direct te starten via de wizard. Data blijft in Europa.

This model is a fine-tuned version of the openai/whisper-large-v3-turbo model finetuned for automatic speech recognition (ASR) in several Kenyan languages, including Swahili, Kalenjin, Kikuyu, Luo, Maasai and Somali. Whisper is a transformer-based encoder-decoder model that conve...

microsoft/paza-whisper-large-v3-turbo Op aanvraag
audio->text · microsoft · sovereign EU
Runs on EU infrastructure operated by European companies; US marketplace capacity is never part of this chain. Full chain
0.8B
Parameters
Contextvenster
8GB
Minimale VRAM
POST /api/v1/audio/transcriptions Op aanvraag

Specificaties

Parameters 0.8B
Minimale VRAM 8 GB
Architectuur WhisperForConditionalGeneration (vLLM)
Licentie mit
Modaliteit audio->text
Uitgebracht January 2026
Uitgever microsoft ↗

Prijzen

Gedeelde router · per token
Op aanvraag
Niet beschikbaar op de gedeelde router. Prijs op aanvraag als dedicated GPU-deployment.
Dedicated GPU · per uur
vanaf €1,95 per uur
Eigen vLLM-instance op Europese cloud (8 GB VRAM), per uur afgerekend.

Gedeelde EU-router, pay-per-token, scale-to-zero. Dedicated GPU-deployments worden per uur afgerekend, zie prijzen.

Direct aanroepen

Drop-in vervanger voor OpenAI: wijzig alleen de base-URL en de API-key. Ook het Anthropic-formaat (/v1/messages) wordt ondersteund.

curl https://hostyourai.com/api/v1/audio/transcriptions \
  -H "Authorization: Bearer hyai-..." \
  -F model="microsoft/paza-whisper-large-v3-turbo" \
  -F file="@meeting.mp3"

Veelgestelde vragen

Kan ik paza whisper large v3 turbo in de EU draaien?

Ja. HostYourAI draait paza whisper large v3 turbo op GPU's in Europese datacenters via vLLM. Prompts en outputs verlaten de EU niet en er is geen Amerikaanse cloudprovider in de keten.

Is paza whisper large v3 turbo hosten AVG/GDPR-compliant?

Ja. Alle verwerking vindt plaats binnen de EU, er is een verwerkersovereenkomst (DPA) beschikbaar en de subprocessor-lijst is openbaar. Open-source gewichten betekenen ook: geen training op jouw data.

Wat kost paza whisper large v3 turbo?

paza whisper large v3 turbo heeft meerdere GPU's tegelijk nodig en draait daarom als dedicated deployment. Je betaalt dan per GPU-uur en niet per token. Vertel ons je volume, dan rekenen we het voor je door.

Is de API compatibel met OpenAI?

Ja. Je gebruikt de standaard OpenAI-SDK's met een aangepaste base-URL (https://hostyourai.com/api/v1). Ook de Anthropic Messages API wordt ondersteund als drop-in.

Andere modellen van Microsoft

Fara1.5 27B

Fara1.5-27B is a multimodal computer use agent (CUA) for web browsers, from Microsoft Research AI Frontiers. It observes the browser through screenshots and acts on the user's behalf by emitting structured tool calls — click, type, scroll, visit URL, web search, and so on — to complete tasks end-to-end.

27B 262K context Bekijk model →
Fara1.5 4B

Fara1.5-4B is a multimodal computer use agent (CUA) for web browsers, from Microsoft Research AI Frontiers. It observes the browser through screenshots and acts on the user's behalf by emitting structured tool calls — click, type, scroll, visit URL, web search, and so on — to complete tasks end-to-end.

4.5B 262K context Bekijk model →
GELab Zero 4B preview Sico Evolution

GELab Zero 4B preview Sico Evolution is een multimodaal taalmodel van Microsoft met 4.4B parameters, gehost op Europese GPU's via een OpenAI-compatibele API.

4.4B Bekijk model →
Fara1.5 9B

Fara1.5-9B is a multimodal computer use agent (CUA) for web browsers, from Microsoft Research AI Frontiers. It observes the browser through screenshots and acts on the user's behalf by emitting structured tool calls — click, type, scroll, visit URL, web search, and so on — to complete tasks end-to-end.

9.4B 262K context Bekijk model →
harrier oss v1 27b

harrier-oss-v1 is a family of multilingual text embedding models developed by Microsoft. The models use decoder-only architectures with last-token pooling and L2 normalization to produce dense text embeddings. They can be applied to a wide range of tasks, including but not limited to retrieval, clustering, semantic similarity, classification, bitext mining, and reranking. The models achieve state-of-the-art results on the Multilingual MTEB v2 benchmark as of the release date.

27B 131K context Bekijk model →
harrier oss v1 270m

harrier-oss-v1 is a family of multilingual text embedding models developed by Microsoft. The models use decoder-only architectures with last-token pooling and L2 normalization to produce dense text embeddings. They can be applied to a wide range of tasks, including but not limited to retrieval, clustering, semantic similarity, classification, bitext mining, and reranking. The models achieve state-of-the-art results on the Multilingual MTEB v2 benchmark as of the release date.

0.3B 33K context Bekijk model →

Probeer paza whisper large v3 turbo gratis

Account aanmaken duurt een minuut. Je API-key werkt direct met paza whisper large v3 turbo.

Start gratis