DeepSeek-V4-Pro GA Release

DeepSeek has launched DeepSeek-V4-Pro, a model designed for production-grade agentic workflows and complex reasoning tasks. This release introduces flexible reasoning effort controls and native integration with the OpenAI Responses API to streamline developer adoption.

Enhanced Agent Capabilities and Reasoning Control

DeepSeek-V4-Pro provides significant performance gains for production agent workflows. To optimize the balance between latency and accuracy, the model introduces a flexible "reasoning effort" setting available for both V4-Pro and V4-Flash.

Users can configure the reasoning effort based on the task complexity:

  • Low: Optimized for simple tasks.
  • High: Designed for daily agent workflows.
  • Max: Reserved for the most complex tasks requiring deep reasoning.

API Integration and Ecosystem Support

DeepSeek-V4-Pro is natively compatible with the OpenAI Responses API, which is optimized for Codex integrations. This allows developers to implement the model via a one-click setup process.

Availability and Access

DeepSeek-V4-Pro is available through the following channels:

  • Web and App: Accessible via "Expert Mode".
  • API: Available using existing model names; setup details are provided in the official API documentation.

Updated API Pricing Model

Effective August 16, 2026, at 16:00 UTC, DeepSeek is updating its API pricing structure. The new system introduces peak and off-peak rates to encourage flexible workload scheduling. Off-peak rates are 50% lower than peak rates.

Sources

Related

  • Dispatch
  • Dispatch
  • Dispatch
  • Dispatch