Monday, 21 September 2026 | Mise à jour quotidienne L'intelligence artificielle au service des constructeurs

DeepSeek V4-Pro

DeepSeek V4-Pro — Spécifications

Rédigé par Mustafa Ihsan d’après la documentation officielle du fournisseur · Dernière mise à jour

Développeur DeepSeek
Type LLM (architecture MoE)
Modalité Texte → Texte
Paramètres 1,6 T au total / ~49 milliards actifs (MoE)
Fenêtre de contexte 1 million
Sortie maximale 384 K
Licence MIT (ouverte)
Poids ouverts Oui
Publié 2026-04
Prix de l’entrée $0.435 /1M
Prix de la sortie $0.87 /1M
Fournisseurs d'API DeepSeek, OpenRouter

Exécutez-le localement

VRAM (4 bits) ~800 Go
GPU minimal Serveur multi-GPU (ex. : 8 × H100 80 Go)

Page officielle →

What is DeepSeek V4-Pro?

DeepSeek V4-Pro is the company’s open flagship: a 1.6-trillion-parameter mixture-of-experts
activating around 49B parameters per token, with a 1M-token context window and MIT-licensed
weights. The API speaks both OpenAI and Anthropic request formats, which makes it unusually
easy to drop into an existing codebase — often a base-URL and key change rather than a
rewrite.

Pricing is $0.435 in / $0.87 out per million tokens, which blends to roughly $0.52 — around
a twentieth of frontier pricing for a model that holds its own on general reasoning. That
ratio, not the parameter count, is why V4-Pro matters. The dual-format API compatibility
compounds it: the switching cost that normally protects incumbent providers largely
disappears, so V4-Pro is a genuine option for teams that would otherwise never evaluate a
Chinese lab’s model. Self-hosting is a data-centre exercise — roughly 800 GB of VRAM at
4-bit, so eight H100 80GBs or more — meaning the open licence here buys auditability and
provider portability rather than a realistic on-premises deployment for most organisations.

DeepSeek V4-Pro pricing: API cost per 1M tokens

Entrée (par million de jetons)$0.435
Sortie (par million de jetons)$0.870
Ratio sortie/entrée
Moyenne pondérée (4:1 entrée:sortie)$0.522 par 1 million de jetons

What DeepSeek V4-Pro costs per month

Dépense mensuelle réelle pour un ratio entrée-sortie de 4:1 — le rapport effectivement observé dans une charge de travail type (chat ou RAG).

Charge de travailJetons/moisCoût par mois
Projet secondaire 1 million en entrée / 0,25 million en sortie $0.65
Petite équipe 20 millions en entrée / 5 millions en sortie $13
Production 200 millions en entrée / 50 millions en sortie $131

Calculez vos propres coûts dans le Calculateur de coûts pour les API IA.

Cheaper alternatives to DeepSeek V4-Pro

ModèleCoût combiné par million de dollarsVous économisez
DeepSeek V4-Flash ouverte $0.168 68 % moins cher

Hébergement local ou paiement via API ?

DeepSeek V4-Pro is open-weight, so you can run it yourself. It needs ~800 Go of VRAM at 4-bit (Multi-GPU server (e.g. 8× H100 80GB)). Self-hosting only beats the API once your volume is high enough to keep that hardware busy — the calculateur auto-hébergement vs API calcule le seuil de rentabilité pour votre volume de jetons.

Questions fréquemment posées

How much does DeepSeek V4-Pro cost per 1M tokens?

DeepSeek V4-Pro costs $0.435 per 1M input tokens and $0.870 per 1M output tokens. At a typical 4:1 input-to-output mix that blends to about $0.522 per 1M tokens.

How much does DeepSeek V4-Pro cost per month?

A small-team workload of 20M input and 5M output tokens a month costs about $13 on DeepSeek V4-Pro. A side project (1M in / 0.25M out) costs roughly $0.65.

What is a cheaper alternative to DeepSeek V4-Pro?

DeepSeek V4-Flash is the strongest cheaper option in our database at $0.168 per 1M blended — about 68% less than DeepSeek V4-Pro. It is also open-weight, so self-hosting is an option.

Can I run DeepSeek V4-Pro locally?

Yes. DeepSeek V4-Pro is open-weight and needs about ~800 GB of VRAM at 4-bit quantisation (Multi-GPU server (e.g. 8× H100 80GB)).

Why does DeepSeek V4-Pro charge more for output than input?

Output tokens are generated one at a time and cannot be batched the way a prompt can, so they cost the provider more to serve. DeepSeek V4-Pro charges 2× more for output, which is why prompt-heavy workloads are far cheaper to run than generation-heavy ones.

See every DeepSeek model priced side by side: Tarifs de l'API DeepSeek.

Les prix indiqués correspondent aux tarifs publics officiels pour l’API principale du modèle et sont mis à jour régulièrement à mesure que les fournisseurs les modifient. Les remises liées aux volumes, aux traitements par lots ou aux entrées mises en cache ne sont pas prises en compte. Comparez tous les modèles côte à côte dans le Base de données des modèles IA ou le Classement des grands modèles linguistiques (LLM).

Défiler vers le haut