Agents AI

Update
ai

DeepSeek Ships V4-Pro-0813 as Its Flagship Model Leaves Preview, Doubling Down on Agent Tasks

DeepSeek officially released DeepSeek-V4-Pro-0813 on August 13, moving its flagship model out of preview with sharply improved agent and coding benchmarks, a 1-million-token context window, and new Responses API and Codex-style tool support.

AgentsAI NewsroomAugust 14, 20262 min read

Flagship model exits preview with an agent-first pitch

DeepSeek formally released DeepSeek-V4-Pro-0813 on Thursday, August 13, taking its flagship model out of the preview status it had carried since earlier this summer. The Chinese AI lab's own changelog describes the release as delivering "significantly enhanced agent capabilities," alongside new support for a Responses API and Codex-style tool integration aimed at developers building agentic applications that call tools, execute code and carry out multi-step workflows with less human oversight.

The model supports a context window of up to 1 million tokens and can generate outputs as long as 384,000 tokens, and it can run in either a "thinking" or "non-thinking" mode depending on the task. DeepSeek reported large gains on agent-oriented benchmarks compared with the preview build: a score of 62.7 on the DeepSWE software-engineering benchmark, versus 12.8 for V4-Pro-Preview, alongside strong results on Terminal-Bench 2.1 and NL2Repo. The model is available immediately through DeepSeek's app, web interface and API.

Mixed independent reception

Independent coverage has been more measured than DeepSeek's own benchmark claims. The South China Morning Post reported that V4-Pro-0813 underperforms some rivals on general reasoning benchmarks while showing particular strength in cybersecurity-related tasks, and outlets including VentureBeat noted the release landed alongside "DeepSeek Harness," an open-source coding-agent harness positioned as a rival to tools like Claude Code. DeepSeek's API pricing page has also flagged a price increase for V4-Pro tokens expected later in August, though the company has not published final figures.

Why it matters

The release keeps DeepSeek in the thick of an accelerating race among Chinese labs — including Moonshot's Kimi and others — to ship open and API-accessible models tuned specifically for agentic coding and tool-use workloads, the same territory OpenAI, Anthropic and xAI are contesting with GPT-5.6, Claude and Grok. With V4-Pro-0813, DeepSeek is signaling that its next competitive battleground is less about raw benchmark leadership and more about being a viable, lower-cost backend for the agent harnesses and coding tools that developers are increasingly building on top of frontier models.

AI-assisted reporting, overseen by the AgentsAI team. Spotted an error? Let us know.