DeepSeek V4 Pro Reaches GA with Adjustable Reasoning and Responses API Support

DeepSeek has released the GA version of V4 Pro for its app, web service and API. The 0813 model adds adjustable reasoning, native Responses API support, and a later hosted route on NVIDIA NIM.
AI Neural Narration
48kHz StudioFish Audio Neural Engine · Natural editorial narration
Key Takeaways
- check_circleDeepSeek announced V4 Pro GA on 13 August 2026 and made it available in Expert Mode on app and web, as well as through its API.
- check_circleThe GA version supports low, high and max reasoning effort, plus native OpenAI Responses API support aimed at Codex workflows.
- check_circleAIMI first saw NVIDIA NIM list deepseek-ai/deepseek-v4-pro-0813 at 00:59:11 SAST on 27 August. That is an endpoint observation, not a new maker release.
From the April preview to the 0813 GA build
DeepSeek first released the V4 preview on 24 April 2026. On 13 August, the company announced the general availability build now identified in its documentation as DeepSeek-V4-Pro-0813.
The API model name remains deepseek-v4-pro, so existing callers do not need to switch to a dated model ID. DeepSeek says the update improves agent performance in production, but that is a vendor claim until independent evaluations test the GA build on comparable workloads.
Reasoning control and agent integrations
The GA release adds adjustable reasoning effort to V4 Pro and V4 Flash. DeepSeek maps low effort to simpler tasks, high to everyday agent work, and max to more complex jobs.
DeepSeek also added native support for the OpenAI Responses API and published a Codex setup path. This gives agent tools another integration option beyond the OpenAI-compatible chat completions and Anthropic interfaces already documented for V4.
Where the GA model is available
DeepSeek lists V4 Pro in Expert Mode on its app and website, and through the official API as deepseek-v4-pro. The company documentation now points that stable API alias to the 0813 build.
AIMI currently records the exact 0813 route on OpenRouter, Cloudflare Workers AI and NVIDIA NIM. OpenRouter first appeared at 18:01:03 SAST on 12 August, Cloudflare at 20:43:56 SAST on 14 August, and NVIDIA NIM at 00:59:11 SAST on 27 August. These are provider observations, not additional maker release dates.
NVIDIA NIM adds a later hosted route
NVIDIA's official Build page now lists deepseek-ai/deepseek-v4-pro-0813 as a preview endpoint on its NIM service. The page dates the Build listing to 24 August 2026, points to DeepSeek's model card for the underlying release, and identifies the service as a trial endpoint governed by NVIDIA's terms.
AIMI first observed the route in NVIDIA's models endpoint at 22:59:11 UTC on 26 August, or 00:59:11 SAST on 27 August. That was exactly 12 days, 4 hours, 15 minutes and 15 seconds after the Cloudflare route was first observed. The gap describes hosting availability, not a second DeepSeek release or a response by NVIDIA.
The API pricing schedule is now active
DeepSeek introduced separate peak and off-peak API rates at 16:00 UTC on 16 August 2026. The company says off-peak rates are 50% below peak pricing.
The current pricing page lists the active windows and rates. Teams moving recurring work to V4 Pro should check that page before hard-coding figures into cost models.
The release timeline needs careful labels
The April date belongs to the V4 preview. The 13 August date belongs to DeepSeek's GA announcement and 0813 model build, while the official model card records the weights publication on the same calendar day. NVIDIA's Build page gives its hosted listing a separate 24 August date.
DeepSeek's announcement gives a calendar date but no precise publication time. AIMI therefore keeps announcement, weights, provider metadata and endpoint first-seen times separate. There were no earlier verified releases in this package's 27 August SAST batch, so the NVIDIA addition is not being framed as part of a same-day release race.
Frequently Asked Questions
Is DeepSeek V4 Pro still a preview?
No. DeepSeek announced V4 Pro GA on 13 August 2026. Some provider routes still contain the word preview, which describes that provider label rather than the current maker release status.
What API model name should developers use?
DeepSeek says the API name remains deepseek-v4-pro. Its documentation identifies the current model behind that alias as DeepSeek-V4-Pro-0813.
Which reasoning effort levels are supported?
DeepSeek documents low, high and max effort for V4 Pro and V4 Flash.
What does the NVIDIA NIM listing mean?
NVIDIA lists deepseek-ai/deepseek-v4-pro-0813 as a preview endpoint. Its 26 August endpoint observation confirms hosted availability at that time, but it is not a new DeepSeek maker release.
When did the new API pricing start?
The peak and off-peak schedule took effect at 16:00 UTC on 16 August 2026. DeepSeek says off-peak rates are 50% lower than peak rates.
Related Frontier Models & Releases
Explore verified specifications, benchmark results, and route pricing across alternative models in this class.
DeepSeek releases V4.1 Flash with native multimodal vision, 552B MoE and lower API rates
DeepSeek has launched DeepSeek-V4.1-Flash, a 552B-parameter mixture-of-experts model featuring a novel Causal Encoder–Decoder architecture with just 8B active input and 16B active output parameters. The release brings native vision, compresses KV cache storage by up to 8x, and slashes off-peak API pricing to $0.15 per million input tokens.
DeepSeek Releases V4 Flash Vision Exp for Multimodal Agent Work
DeepSeek has released V4 Flash Vision Exp, an experimental API model that adds image understanding to the V4 Flash family while keeping its text capabilities.
OpenAI Releases GPT-6 Astra: Next-Generation Flagship with 1.05M Context and Deep Multimodal Reasoning
OpenAI has officially launched GPT-6 Astra, its frontier flagship model featuring a 1,050,000-token context window, 128,000 max output tokens, and native tool-use for autonomous agent workflows.
Primary Sources
- DeepSeek V4 Pro GA announcement
- DeepSeek V4 Pro model card and weights
- NVIDIA NIM DeepSeek-V4-Pro-0813 model page
- NVIDIA NIM models endpoint
- DeepSeek API changelog
- DeepSeek models and pricing
- DeepSeek thinking mode guide
- DeepSeek Codex integration guide
- OpenRouter DeepSeek V4 Pro route
- Cloudflare Workers AI model catalogue