Recevez 100 crédits gratuits à l'inscription pour explorer et créer vos applications d'IAObtenir
deepseek-v4-pro

DeepSeek V4 Pro / Chat

Commercial
ID: deepseek-v4-pro

Large-scale MoE model with 1.6T total parameters and 49B activated, supporting a 1M-token context window . Designed for advanced reasoning, coding, and long-horizon agent workflows . Top performance on GPQA Diamond (88.8%) and Terminal-Bench Hard (46.2%)

Entrée$0.66/ 1M Tokens(660 Crédits)
Sortie$1.98/ 1M Tokens(1980 Crédits)

Press Enter to add, Backspace to remove

1
1
Conversation Chat

Playground Chat

Démarrez une conversation avec le modèle IA. Vous pouvez poser toutes vos questions.

0

Les réponses générées par l'IA peuvent varier en précision.

Détails des tarifs

La facturation réelle de ce modèle est calculée dynamiquement en fonction des paramètres spécifiques de votre requête API. Voici les combinaisons spécifiques et leurs tarifs correspondants :

Remarque
  • Pour les requêtes effectuées en semaine (lun-ven) durant 09:00 - 12:00, 14:00 - 18:00 (UTC+8), un multiplicateur de prix de 2x est appliqué.
Standard
Prix d'entrée
$0.660(660 / 1M Tokens)
Prix de sortie
$1.980(1,980 / 1M Tokens)
Cache implicite
$0.022(22/ 1M Tokens)
Cache explicite
$0.022(22/ 1M Tokens)
Création de cache
--
Ecosystem Models

Recommended Related Models

Explore complementary video and multimodal models with your unified API key.

Browse All Models
GLM-5.3
chatCommercial
GLM-5.3

GLM-5.3

glm-5.3

Zhipu AI's flagship text model optimized for complex software engineering and long-horizon Agent tasks. Features a 1M-token context window with mandatory reasoning (3 levels: low/high/max). Coding capability improved 50% over GLM-5.2 on Z.ai Code Bench; scores SOTA on Terminal Bench 3.0. Excels in cybersecurity tasks (cyber vulnerability discovery)

ChatText Generation
Qwen3.8 Flash
chatCommercial
Qwen3.8 Flash

Qwen3.8 Flash

qwen3.8-flash

Alibaba's cost-efficient multimodal reasoning model. Supports text, image, and video inputs with text output. Features a native 1M-token context window for long documents, codebases, and agentic workflows. Excels in coding assistance, desktop interaction, chart analysis, and long-video understanding. Compatible with OpenAI/Anthropic protocols for seamless integration

ChatText Generation
GLM-5.2
chatCommercial
GLM-5.2

GLM-5.2

glm-5.2

Flagship text model purpose-built for long-horizon agentic workflows. Features a 1M context window supporting project-level engineering in a single session. Excels at autonomous coding: can complete development, testing, and multi-platform deployment from a single prompt. Top open-weight model per Artificial Analysis; #1 globally on Code Arena. MIT-licensed and Day-0 optimized for domestic AI chips

ChatText Generation
Qwen3 Max
chatCommercial
Qwen3 Max

Qwen3 Max

qwen3-max

Compared with the September 23, 2025 version, the newly upgraded Qwen-3 Max seamlessly integrates thinking and non-thinking modes, bringing an all-round obvious performance boost. Its thinking mode supports web search, web content extraction and code interpreter. It can conduct in-depth logical reasoning and call external tools to solve intricate problems more precisely

Chat