DeepSeek released DeepSeek-V4-Pro for general use, following its preview in April, with improvements aimed at software agents that carry out multi-step tasks. V4-Pro and V4-Flash now let callers pick how much reasoning effort a request uses, from low to maximum, and support OpenAI's Responses API directly. New API pricing with cheaper off-peak rates takes effect on 16 August 2026; existing model names are unchanged.
What changed
V4 was available only as a preview, with flat API pricing and no adjustable reasoning setting.
What it unlocks
Choosing how much reasoning effort a request uses, and scheduling heavy API workloads into cheaper off-peak hours.
- off-peak 50% below peak rates
- new pricing from 16:00 UTC, 16 Aug 2026
What you need to act on it
- API access, or the app/web "Expert Mode" for the chat version
- api-docs.deepseek.com2026-08-13