[DeepSeek](/ai-gateway/models/labs/deepseek)

# DeepSeek V4 Pro

DeepSeek V4 Pro is DeepSeek's April 23, 2026 top-tier model in the V4 series. It pairs a hybrid attention architecture with a context window of 1.0M tokens and targets complex reasoning, multi-step problem solving, and agentic tasks. Your use is subject to DeepSeek's [Terms](https://cdn.deepseek.com/policies/en-US/deepseek-terms-of-use.html) & [Privacy](https://cdn.deepseek.com/policies/en-US/deepseek-privacy-policy.html) Policies.

ReasoningTool UseImplicit Caching

[Use with AI Gateway](https://vercel.com/d?to=%2F%5Bteam%5D%2F%7E%2Fai%3Futm_source%3Dgateway-model-page%26utm_campaign%3Dai-gateway-models&title=Get+Started+with+Vercel+AI+Gateway) [View docs](https://vercel.com/docs/ai-gateway)

AI SDKChat CompletionsMessagesResponses

```
1import { streamText } from 'ai'
2

3const result = streamText({
4  model: 'deepseek/deepseek-v4-pro',
5  prompt: 'Why is the sky blue?'
6})
```

[Read docs](https://vercel.com/docs/ai-gateway/sdks-and-apis/ai-sdk)

[Overview](/ai-gateway/models/deepseek-v4-pro) [API](/ai-gateway/models/deepseek-v4-pro/api) [About](/ai-gateway/models/deepseek-v4-pro/about) [Providers](/ai-gateway/models/deepseek-v4-pro/providers) [Throughput](/ai-gateway/models/deepseek-v4-pro/throughput) [Latency](/ai-gateway/models/deepseek-v4-pro/latency) [Uptime](/ai-gateway/models/deepseek-v4-pro/uptime) [Status](/ai-gateway/models/deepseek-v4-pro/status) [Similar](/ai-gateway/models/deepseek-v4-pro/similar) [FAQ](/ai-gateway/models/deepseek-v4-pro/faq)

## [Copy link to heading](#playground)Playground

Try out DeepSeek V4 Pro by DeepSeek. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.

![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75)DeepSeek V4 Pro

![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=96&q=75)

DeepSeek V4 Pro

## [Copy link to heading](#providers)Providers

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the [docs](/docs/ai-gateway/provider-options) for more info. Using a provider means you agree to their terms, listed under Legal.

Token prices are 2× during 01:00–04:00 and 06:00–10:00 UTC. Listed prices are the off-peak rate.

| Provider |
| --- |

| Context | Max Output | Latency | Throughput | Input | Output | Cache | Web Search | Capabilities | ZDR | No Training | Regional Inference | Release Date |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

| ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) [DeepSeek](/ai-gateway/models/providers/deepseek) Legal:[Terms](https://cdn.deepseek.com/policies/en-US/deepseek-terms-of-use.html)•[Privacy](https://cdn.deepseek.com/policies/en-US/deepseek-privacy-policy.html) | 1M | 384K | 1.8s | 65tps | $0.66/M | $1.98/M | Read:$0.02/M Write:— | — |  |  |  |  | 04/23/2026 |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |
| ![fireworks logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Ffireworks.png&w=48&q=75) Going away Aug 27, 2026 [Fireworks](/ai-gateway/models/providers/fireworks) Legal:[Terms](https://fireworks.ai/terms-of-service)•[Privacy](https://fireworks.ai/privacy-policy) | 1M | 1M | 1.3s | 69tps | $1.74/M | $3.48/M | Read:$0.14/M Write:— | — |  |  |  |  | 04/23/2026 |  |
| ![novita logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fnovita.png&w=48&q=75) [Novita AI](/ai-gateway/models/providers/novita) Legal:[Terms](https://novita.ai/legal/terms-of-service)•[Privacy](https://novita.ai/legal/privacy-policy) | 1M | 393K | 1.2s | 66tps | $1.74/M | $3.48/M | Read:$0.14/M Write:— | — |  |  |  |  | 04/23/2026 |  |
| ![deepinfra logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepinfra.png&w=48&q=75) [DeepInfra](/ai-gateway/models/providers/deepinfra) Legal:[Terms](https://deepinfra.com/terms)•[Privacy](https://deepinfra.com/privacy) | 66K | 66K | 1.1s | 61tps | $1.74/M | $3.48/M | Read:$0.14/M Write:— | — |  |  |  |  | 04/23/2026 |  |
| ![baseten logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fbaseten.png&w=48&q=75) [Baseten](/ai-gateway/models/providers/baseten) Legal:[Terms](https://www.baseten.co/terms-and-conditions/)•[Privacy](https://www.baseten.co/privacy-policy/) | 1M | 1M | 0.7s | 160tps | $1.74/M | $3.48/M | Read:$0.15/M Write:— | — |  |  |  | US | 04/23/2026 |  |
| ![togetherai logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Ftogetherai.png&w=48&q=75) [Together AI](/ai-gateway/models/providers/togetherai) Legal:[Terms](https://www.together.ai/terms-of-service)•[Privacy](https://www.together.ai/privacy) | 1M | 1M | 0.5s | 81tps | $2.10/M | $4.40/M | Read:$0.2/M Write:— | — |  |  |  |  | 04/23/2026 |  |
| ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) [Azure](/ai-gateway/models/providers/azure) Legal:[Terms](https://learn.microsoft.com/en-us/legal/cognitive-services/openai/code-of-conduct)•[Privacy](https://privacy.microsoft.com/en-us/privacystatement) | 1M | 128K | 0.7s | 96tps | $1.74/M | $3.48/M | Read:$0.14/M Write:— | — |  |  |  |  | 04/23/2026 |  |
| ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75) [Alibaba Cloud](/ai-gateway/models/providers/alibaba) Legal:[Terms](https://www.alibabacloud.com/help/en/legal/latest/alibaba-cloud-international-website-product-terms-of-service-v-3-8-0)•[Privacy](https://www.alibabacloud.com/help/en/legal/latest/alibaba-cloud-international-website-privacy-policy) | 1M | 384K | 3.1s | 52tps | $1.65/M | $3.30/M | Read:$0.14/M Write:— | — |  |  |  |  | 04/23/2026 |  |

## [Copy link to heading](#throughput)Throughput24 hours

1W

1D

P50 throughput on live AI Gateway traffic, in tokens per second (TPS). Visit the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/metrics) for more info.

## [Copy link to heading](#latency)Latency24 hours

1W

1D

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/metrics) for more info.

## [Copy link to heading](#uptime)Uptime24 hours

1W

1D

1H

Direct request success rate on AI Gateway and per-provider. Visit the [docs](https://vercel.com/docs/ai-gateway/models-and-providers/uptime) for more info.

1W

1D

1H

## [Copy link to heading](#more-models-by-deepseek)More models by DeepSeek

All

Text

Code

| Model |
| --- |

| Context | Latency | Throughput | Input | Output | Cache | Web Search | Capabilities | Providers | ZDR | No Training | Release Date |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |

| ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) [deepseek/deepseek-v4-flash-vision-exp](/ai-gateway/models/deepseek-v4-flash-vision-exp) | 1M | 1.4s | 125tps | $0.22/M | $0.66/M | Read:$0.01/M Write:— | — | +1 | ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) |  |  | 08/21/2026 |  |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- |
| ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) [deepseek/deepseek-v4-pro-0813](/ai-gateway/models/deepseek-v4-pro-0813) | 1M | 0.6s | 116tps | $0.66/M | $1.98/M | Read:$0.02/M Write:— | — |  | ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75) ![baseten logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fbaseten.png&w=48&q=75) ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) +2 |  |  | 08/12/2026 |  |
| ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) [deepseek/deepseek-v4-flash-0731](/ai-gateway/models/deepseek-v4-flash-0731) | 1M | 0.4s | 259tps | $0.08/M | $0.15/M | Read:$0.01/M Write:— | — |  | ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75) ![baseten logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fbaseten.png&w=48&q=75) ![deepinfra logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepinfra.png&w=48&q=75) +7 |  |  | 07/31/2026 |  |
| ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) [deepseek/deepseek-v4-flash](/ai-gateway/models/deepseek-v4-flash) | 1M | 0.9s | 209tps | $0.09/M | $0.18/M | Read:$0.01/M Write:— | — |  | ![alibaba logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Falibaba%2520cloud.png&w=48&q=75) ![azure logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fazure.png&w=48&q=75) ![baseten logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fbaseten.png&w=48&q=75) +7 |  |  | 04/23/2026 |  |
| ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) [deepseek/deepseek-v3.2](/ai-gateway/models/deepseek-v3.2) | 164K | 0.6s | 98tps | $0.28/M | $0.42/M | Read:$0.03/M Write:— | — |  | ![bedrock logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Famazon%2520bedrock.png&w=48&q=75) ![deepinfra logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepinfra.png&w=48&q=75) ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) +1 |  |  | 12/01/2025 |  |
| ![deepseek logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepseek.png&w=48&q=75) [deepseek/deepseek-v3.2-thinking](/ai-gateway/models/deepseek-v3.2-thinking) | 164K | 0.5s | 45tps | $0.26/M | $0.38/M | Read:$0.13/M Write:— | — |  | ![bedrock logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Famazon%2520bedrock.png&w=48&q=75) ![deepinfra logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fdeepinfra.png&w=48&q=75) ![novita logo](/vc-ap-vercel-marketing/_next/image?url=https%3A%2F%2F7nyt0uhk7sse4zvn.public.blob.vercel-storage.com%2Fdocs-assets%2Fstatic%2Fdocs%2Fai-gateway%2Flogos%2Fnovita.png&w=48&q=75) |  |  | 12/01/2025 |  |

## [Copy link to heading](#about-deepseek-v4-pro)About DeepSeek V4 Pro

DeepSeek V4 Pro was released April 23, 2026 as the high-capability tier of DeepSeek's V4 generation. The V4 series introduces a hybrid attention architecture that combines Compressed Sparse Attention (CSA) with Heavily Compressed Attention (HCA), and uses ManifoldConstrained Hyper-Connections (mHC) in place of standard residual connections. The combination supports efficient inference at the 1.0M tokens window.

DeepSeek V4 Pro is positioned for complex reasoning, multi-step problem solving, and agentic workflows. Tool use, reasoning, and implicit caching are all supported, so DeepSeek V4 Pro fits planner-style pipelines where the model decides on tool calls, integrates results, and iterates toward an answer. Maximum output is 1.0M tokens, which gives long-form reasoning chains and tool-call sequences room to complete in a single response.

Access is through AI Gateway with an AI Gateway API key or OIDC token. You can integrate through the AI SDK, Chat Completions, Responses, or Messages API formats. Implicit caching applies when a long input prefix repeats across calls, charging the cached input rate instead of the standard input rate for cached tokens.

## [Copy link to heading](#what-to-consider-when-choosing-a-provider)What To Consider When Choosing a Provider

- Configuration: DeepSeek V4 Pro is priced higher than the V4 Flash variant. If your workload is short-form instruction following or classification, DeepSeek V4 Flash delivers similar output limits at lower per-token rates.
- Zero Data Retention: Zero Data Retention is available for this model. It is offered on a per-provider and model basis. See the [documentation](https://vercel.com/docs/ai-gateway/security-and-compliance/zdr) for details.
- Authentication: AI Gateway authenticates requests using an [API key](https://vercel.com/docs/ai-gateway/authentication-and-byok#api-key-authentication) or [OIDC token](https://vercel.com/docs/ai-gateway/authentication-and-byok#oidc-token-authentication). You do not need to manage provider credentials directly.

## [Copy link to heading](#when-to-use-deepseek-v4-pro)When to Use DeepSeek V4 Pro

### Best for

- Complex reasoning workloads: Analytical research, technical synthesis, and structured derivation benefit from the 1.0M tokens window and V4 architecture
- Multi-step agent pipelines: Planning, tool calls, and result integration combine in a single endpoint
- Software engineering automation: Reasoning and reliable tool use across long context support code-generation pipelines
- Multi-format API integrations: AI SDK, Chat Completions, Responses, and Messages API formats route to the same high-capability model

### Consider alternatives when

- Short-form throughput: Use DeepSeek V4 Flash for classification, routing, and instruction following at lower per-token cost
- MIT-licensed reasoning: DeepSeek-R1 remains the open-weights reasoning specialist when license terms drive the selection
- Earlier-generation cost: DeepSeek V3 family models may meet capability needs at lower cost when V4 context and architecture aren't required

## [Copy link to heading](#conclusion)Conclusion

DeepSeek V4 Pro is the capability tier of the V4 generation, suited to complex reasoning and agentic workloads at the 1.0M tokens window. For short-form, high-volume tasks within the same generation, DeepSeek V4 Flash is the cost-efficient alternative.