AI Research
DeepSeek V4 Pro Reaches GA with Adjustable Reasoning and Responses API Support
DeepSeek has released the GA version of V4 Pro for its app, web service and API. The 0813 model adds adjustable reasoning, native Responses API support, and a later hosted route on NVIDIA NIM.
Route DeepSeek V4 Pro through AZ Labs AI GatewayAI-narrated summary, not a word-for-word reading.

Key takeaways
- check_circleDeepSeek announced V4 Pro GA on 13 August 2026 and made it available in Expert Mode on app and web, as well as through its API.
- check_circleThe GA version supports low, high and max reasoning effort, plus native OpenAI Responses API support aimed at Codex workflows.
- check_circleAIMI first saw NVIDIA NIM list deepseek-ai/deepseek-v4-pro-0813 at 00:59:11 SAST on 27 August. That is an endpoint observation, not a new maker release.
From the April preview to the 0813 GA build
DeepSeek first released the V4 preview on 24 April 2026. On 13 August, the company announced the general availability build now identified in its documentation as DeepSeek-V4-Pro-0813.
The API model name remains deepseek-v4-pro, so existing callers do not need to switch to a dated model ID. DeepSeek says the update improves agent performance in production, but that is a vendor claim until independent evaluations test the GA build on comparable workloads.
Reasoning control and agent integrations
The GA release adds adjustable reasoning effort to V4 Pro and V4 Flash. DeepSeek maps low effort to simpler tasks, high to everyday agent work, and max to more complex jobs.
DeepSeek also added native support for the OpenAI Responses API and published a Codex setup path. This gives agent tools another integration option beyond the OpenAI-compatible chat completions and Anthropic interfaces already documented for V4.
Where the GA model is available
DeepSeek lists V4 Pro in Expert Mode on its app and website, and through the official API as deepseek-v4-pro. The company documentation now points that stable API alias to the 0813 build.
AIMI currently records the exact 0813 route on OpenRouter, Cloudflare Workers AI and NVIDIA NIM. OpenRouter first appeared at 18:01:03 SAST on 12 August, Cloudflare at 20:43:56 SAST on 14 August, and NVIDIA NIM at 00:59:11 SAST on 27 August. These are provider observations, not additional maker release dates.
NVIDIA NIM adds a later hosted route
NVIDIA's official Build page now lists deepseek-ai/deepseek-v4-pro-0813 as a preview endpoint on its NIM service. The page dates the Build listing to 24 August 2026, points to DeepSeek's model card for the underlying release, and identifies the service as a trial endpoint governed by NVIDIA's terms.
AIMI first observed the route in NVIDIA's models endpoint at 22:59:11 UTC on 26 August, or 00:59:11 SAST on 27 August. That was exactly 12 days, 4 hours, 15 minutes and 15 seconds after the Cloudflare route was first observed. The gap describes hosting availability, not a second DeepSeek release or a response by NVIDIA.
The API pricing schedule is now active
DeepSeek introduced separate peak and off-peak API rates at 16:00 UTC on 16 August 2026. The company says off-peak rates are 50% below peak pricing.
The current pricing page lists the active windows and rates. Teams moving recurring work to V4 Pro should check that page before hard-coding figures into cost models.
The release timeline needs careful labels
The April date belongs to the V4 preview. The 13 August date belongs to DeepSeek's GA announcement and 0813 model build, while the official model card records the weights publication on the same calendar day. NVIDIA's Build page gives its hosted listing a separate 24 August date.
DeepSeek's announcement gives a calendar date but no precise publication time. AIMI therefore keeps announcement, weights, provider metadata and endpoint first-seen times separate. There were no earlier verified releases in this package's 27 August SAST batch, so the NVIDIA addition is not being framed as part of a same-day release race.
Frequently asked questions
Is DeepSeek V4 Pro still a preview?
No. DeepSeek announced V4 Pro GA on 13 August 2026. Some provider routes still contain the word preview, which describes that provider label rather than the current maker release status.
What API model name should developers use?
DeepSeek says the API name remains deepseek-v4-pro. Its documentation identifies the current model behind that alias as DeepSeek-V4-Pro-0813.
Which reasoning effort levels are supported?
DeepSeek documents low, high and max effort for V4 Pro and V4 Flash.
What does the NVIDIA NIM listing mean?
NVIDIA lists deepseek-ai/deepseek-v4-pro-0813 as a preview endpoint. Its 26 August endpoint observation confirms hosted availability at that time, but it is not a new DeepSeek maker release.
When did the new API pricing start?
The peak and off-peak schedule took effect at 16:00 UTC on 16 August 2026. DeepSeek says off-peak rates are 50% lower than peak rates.
Sources
- DeepSeek V4 Pro GA announcement
- DeepSeek V4 Pro model card and weights
- NVIDIA NIM DeepSeek-V4-Pro-0813 model page
- NVIDIA NIM models endpoint
- DeepSeek API changelog
- DeepSeek models and pricing
- DeepSeek thinking mode guide
- DeepSeek Codex integration guide
- OpenRouter DeepSeek V4 Pro route
- Cloudflare Workers AI model catalogue
Copyright & image credits
The article image is credited to © DeepSeek, used for news reporting. Source: DeepSeek: DeepSeek-V4-Pro GA Release.