Model garden Router · auf Anfrage Dedicated · verfügbar

DeepSeek OCR 2

Läuft als eigenes dediziertes GPU-Deployment, direkt über den Wizard startbar. Daten bleiben in Europa.

Inference using Huggingface transformers on NVIDIA GPUs. Requirements tested on python 3.12.9 + CUDA11.8:

deepseek-ai/DeepSeek-OCR-2 vLLM ready
text+image->text · deepseek-ai · sovereign EU
Runs on EU infrastructure operated by European companies; US marketplace capacity is never part of this chain. Full chain
3.4B
Parameter
8K
Kontextfenster
16GB
Minimaler VRAM
POST /api/v1/chat/completions Auf Anfrage

Spezifikationen

Parameter 3.4B
Kontextfenster 8,192 tokens
Minimaler VRAM 16 GB
Architektur DeepseekOCR2ForCausalLM (vLLM)
Lizenz apache-2.0
Modalität text+image->text
Veröffentlicht January 2026
Anbieter deepseek-ai ↗

Preise

Gemeinsamer Router · pro Token
Auf Anfrage
Nicht über den gemeinsamen Router verfügbar. Preis auf Anfrage als dediziertes GPU-Deployment.
Dediziertes GPU · pro Stunde
ab €1,68 pro Stunde
Eigene vLLM-Instanz in europäischer Cloud (16 GB VRAM), stundenweise abgerechnet.

Gemeinsamer EU-Router, Pay-per-Token, Scale-to-Zero. Dedizierte GPU-Deployments werden stundenweise abgerechnet, siehe Preise.

✓ Funktionsfähig verifiziert am 27-07-2026, Antwort in 438 ms auf unserer EU-Infrastruktur.

Direkt aufrufen

Drop-in-Ersatz für OpenAI: Ändern Sie nur die Base-URL und den API-Key. Auch das Anthropic-Format (/v1/messages) wird unterstützt.

curl https://hostyourai.com/api/v1/chat/completions \
  -H "Authorization: Bearer hyai-..." \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-ai/DeepSeek-OCR-2",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Häufig gestellte Fragen

Kann ich DeepSeek OCR 2 in der EU betreiben?

Ja. HostYourAI betreibt DeepSeek OCR 2 auf GPUs in europäischen Rechenzentren über vLLM. Prompts und Outputs verlassen die EU nicht und es ist kein US-Cloud-Anbieter in der Kette.

Ist das Hosting von DeepSeek OCR 2 DSGVO-konform?

Ja. Die gesamte Verarbeitung findet innerhalb der EU statt, ein Auftragsverarbeitungsvertrag (AVV) ist verfügbar und die Liste der Subprozessoren ist öffentlich. Open-Source-Gewichte bedeuten außerdem: kein Training mit Ihren Daten.

Was kostet DeepSeek OCR 2?

DeepSeek OCR 2 ist auf Anfrage verfügbar: über den gemeinsamen EU-Router nach kurzer Abstimmung oder als dediziertes GPU-Deployment, das pro Stunde abgerechnet wird. Nennen Sie uns Ihren Anwendungsfall, dann richten wir es für Sie ein.

Ist die API OpenAI-kompatibel?

Ja. Sie verwenden die Standard-OpenAI-SDKs mit einer angepassten Base-URL (https://hostyourai.com/api/v1). Auch die Anthropic Messages API wird als Drop-in unterstützt.

Weitere Modelle von DeepSeek

DeepSeek V4 Pro 0813

DeepSeek-V4-Pro-0813 is the official release of DeepSeek-V4-Pro, superseding the preview version, with greatly enhanced agentic capabilities and performance improvements that are especially pronounced in production environments. It is built on the DeepSeek-V4-Pro (Preview) model structure, with a DSpark speculative decoding module attached.

1650B 1M Kontext Modell ansehen →
DeepSeek V4 Pro

We present a preview version of DeepSeek-V4 series, including two strong Mixture-of-Experts (MoE) language models — DeepSeek-V4-Pro with 1.6T parameters (49B activated) and DeepSeek-V4-Flash with 284B parameters (13B activated) — both supporting a context length of one million tokens.

1M Kontext Modell ansehen →
DeepSeek V4 Flash

We present a preview version of DeepSeek-V4 series, including two strong Mixture-of-Experts (MoE) language models — DeepSeek-V4-Pro with 1.6T parameters (49B activated) and DeepSeek-V4-Flash with 284B parameters (13B activated) — both supporting a context length of one million tokens.

291B 1M Kontext Modell ansehen →
DeepSeek V3.2

We introduce DeepSeek-V3.2, a model that harmonizes high computational efficiency with superior reasoning and agent performance. Our approach is built upon three key technical breakthroughs:

164K Kontext Modell ansehen →
DeepSeek OCR

torch==2.6.0 transformers==4.46.3 tokenizers==0.20.3 einops addict easydict pip install flash-attn==2.7.3 --no-build-isolation

3.3B 8K Kontext Modell ansehen →
DeepSeek V3.2 Exp

We are excited to announce the official release of DeepSeek-V3.2-Exp, an experimental version of our model. As an intermediate step toward our next-generation architecture, V3.2-Exp builds upon V3.1-Terminus by introducing DeepSeek Sparse Attention—a sparse attention mechanism designed to explore and validate optimizations for training and inference efficiency in long-context scenarios.

164K Kontext Modell ansehen →

Testen Sie DeepSeek OCR 2 kostenlos

Die Kontoerstellung dauert eine Minute. Testen Sie DeepSeek OCR 2 direkt im Playground.

Kostenlos starten