DeepSeek Ships V4-Pro-0813 as Its Flagship Model Leaves Preview, Doubling Down on Agent Tasks
DeepSeek officially released DeepSeek-V4-Pro-0813 on August 13, moving its flagship model out of preview with sharply improved agent and coding benchmarks, a 1-million-token context window, and new Responses API and Codex-style tool support.
Flagship model exits preview with an agent-first pitch
DeepSeek formally released DeepSeek-V4-Pro-0813 on Thursday, August 13, taking its flagship model out of the preview status it had carried since earlier this summer. The Chinese AI lab's own changelog describes the release as delivering "significantly enhanced agent capabilities," alongside new support for a Responses API and Codex-style tool integration aimed at developers building agentic applications that call tools, execute code and carry out multi-step workflows with less human oversight.
The model supports a context window of up to 1 million tokens and can generate outputs as long as 384,000 tokens, and it can run in either a "thinking" or "non-thinking" mode depending on the task. DeepSeek reported large gains on agent-oriented benchmarks compared with the preview build: a score of 62.7 on the DeepSWE software-engineering benchmark, versus 12.8 for V4-Pro-Preview, alongside strong results on Terminal-Bench 2.1 and NL2Repo. The model is available immediately through DeepSeek's app, web interface and API.
Mixed independent reception
Independent coverage has been more measured than DeepSeek's own benchmark claims. The South China Morning Post reported that V4-Pro-0813 underperforms some rivals on general reasoning benchmarks while showing particular strength in cybersecurity-related tasks, and outlets including VentureBeat noted the release landed alongside "DeepSeek Harness," an open-source coding-agent harness positioned as a rival to tools like Claude Code. DeepSeek's API pricing page has also flagged a price increase for V4-Pro tokens expected later in August, though the company has not published final figures.
Why it matters
The release keeps DeepSeek in the thick of an accelerating race among Chinese labs — including Moonshot's Kimi and others — to ship open and API-accessible models tuned specifically for agentic coding and tool-use workloads, the same territory OpenAI, Anthropic and xAI are contesting with GPT-5.6, Claude and Grok. With V4-Pro-0813, DeepSeek is signaling that its next competitive battleground is less about raw benchmark leadership and more about being a viable, lower-cost backend for the agent harnesses and coding tools that developers are increasingly building on top of frontier models.
Sources
- DeepSeek officially launches V4-Pro AI model in August 2026 — Yahoo Tech
- DeepSeek's updated V4 Pro AI model struggles on benchmarks, shines in cybersecurity — South China Morning Post
- DeepSeek Harness launches as open source rival to Claude Code, alongside V4-Pro on API — VentureBeat
- Change Log — DeepSeek API Docs
AI-assisted reporting, overseen by the AgentsAI team. Spotted an error? Let us know.
Related agents
More ai news
EU Set to Propose Barring Under-15s From AI Chatbots and Social Media in 'Kids Act'
The European Commission is preparing to unveil an EU Kids Act that would bar unsupervised access to AI chatbots, social media, video platforms and online games for under-15s, with tiered rules and mandatory age verification for 13-14 year-olds.
Amodei's 'Pace the Frontier' Plan Draws Same-Day Backing From OpenAI, DeepMind and xAI
Anthropic CEO Dario Amodei published an essay arguing frontier AI labs should deliberately slow capability gains, and within hours Sam Altman, Demis Hassabis and Elon Musk publicly endorsed the idea, with Microsoft's Satya Nadella following a day later.
Positron Raises $875M to Build an HBM-Free AI Inference Chip
Chip startup Positron closed an $875 million Series C at a $5 billion post-money valuation to fund its Asimov inference accelerator, which pairs its compute architecture with up to 2,304GB of commodity LPDDR5X memory instead of scarce high-bandwidth memory.
DeepSeek Releases V4.1 Flash, Cuts API Prices and Sets End Date for V4 Pro
DeepSeek officially released V4.1 Flash, a cheaper and faster multimodal successor to V4 Pro with a 1-million-token context window, and said it will reroute all V4 Pro API traffic to the new model from September 14.