# DeepSeek-V4-Pro-0813

DeepSeek-V4-Pro-0813 is an AI model released by DeepSeek on Aug 13 2026. It has open weight. Tracked results include 35% on BullshitBench v2, 80.6% on SWE-Bench Verified and 60% on Humanity's Last Exam (with tools).

## Facts

| Field | Value |
| --- | --- |
| Model | DeepSeek-V4-Pro-0813 |
| Developer | DeepSeek |
| Release date | Thursday, Aug 13 2026 |
| Licensing | Open Weight |

## API pricing

All rates in USD per 1,000,000 tokens, pay-as-you-go.

| Tier | Input | Cached input | Output |
| --- | --- | --- | --- |
| Off-peak (all hours outside the published peak windows) | $0.66 | $0.022 | $1.98 |
| Peak (01:00–04:00 and 06:00–10:00 UTC) | $1.32 | $0.044 | $3.96 |

Verified August 18, 2026 against the first-party source: https://api-docs.deepseek.com/quick_start/pricing

## Tracked benchmark scores

| Benchmark | Score | Source | What it measures |
| --- | --- | --- | --- |
| BullshitBench v2 | 35% | [BullshitBench](https://github.com/petergpt/bullshit-benchmark) | Given a confidently-worded but nonsensical prompt, does the AI spot that it makes no sense and push back — instead of playing along and inventing an answer? The score is how often it clearly called out the nonsense. Higher is better. |
| SWE-Bench Verified | 80.6% | [BenchLM](https://benchlm.ai), retrieved 2026-08-31 | Real coding tasks pulled from open-source projects — the AI has to find and fix actual bugs. A human-checked version of the original SWE-Bench. Higher is better. |
| Terminal-Bench 2.1 | 87.9% | Lab | Can the AI work in a command-line terminal — running commands and finishing technical setup tasks the way a developer would? Higher is better. |
| Toolathlon-Verified | 74.1% | Lab | Tests how well the AI uses everyday personal tools and apps to get things done — a human-checked version of Toolathlon. Higher is better. |
| BrowseComp | 83.4% | [BenchLM](https://benchlm.ai), retrieved 2026-08-31 | Can the AI browse the web and track down hard-to-find answers? Higher is better. |
| CyberGym | 83.3% | Lab | Tests the AI on cybersecurity challenges — finding and exploiting software weaknesses inside a safe sandbox. Higher is better. |
| Humanity's Last Exam (no tools) | 42.7% | Lab | Humanity's Last Exam — extremely hard expert questions across many subjects, written so you can't just look up the answer. “No tools” means the AI answers on its own. Higher is better. |
| Humanity's Last Exam (with tools) | 60% | Lab | Humanity's Last Exam — extremely hard expert questions across many subjects. “With tools” means the AI is allowed to search the web or run code while answering. Higher is better. |
| AutomationBench | 31.8% | Lab | Tests whether the AI can run real multi-step business workflows — the kind of end-to-end office processes companies want to automate — from start to finish. Higher is better. |

## About DeepSeek-V4-Pro-0813

DeepSeek-V4-Pro-0813, released August 13, 2026, brought the top tier of the V4 line up to the agentic post-training that had landed on Flash two weeks earlier. On DeepSeek's own harness it scored 87.9 on Terminal-Bench 2.1, against 82.7 for V4-Flash-0731 and 72.1 for the April V4-Pro-Preview build, and 83.3 on CyberGym — the highest figure in the lab's launch comparison at the time, marginally ahead of Claude Fable 5. It also posted 42.7% on Humanity's Last Exam without tools and 60.0% with them, 74.1 on Toolathlon-Verified and 31.8 on the public AutomationBench split. As with the July Flash release, DeepSeek ran the code-agent evaluations through its own unreleased harness in minimal mode, and its numbers for rival models did not always match those labs' published figures, so the chart reads best as a within-family comparison.

The headline changes were operational rather than architectural. V4-Pro and V4-Flash both gained a selectable reasoning effort — low for simple prompts, high for everyday agent work, max for hard tasks — replacing the single fixed thinking budget of the earlier builds, and the release added native OpenAI Responses API support with one-click Codex setup. Distribution followed the pattern the Flash refresh established: the checkpoint reached the app and web product under an "Expert Mode" toggle and shipped through the API under unchanged model names, so existing callers were moved onto it without a code change. Weights for this checkpoint are available under the MIT License on Hugging Face.

## Questions and answers

### When was DeepSeek-V4-Pro-0813 released?

DeepSeek-V4-Pro-0813 was released by DeepSeek on Thursday, Aug 13 2026.

### Who made DeepSeek-V4-Pro-0813?

DeepSeek-V4-Pro-0813 was built by DeepSeek. Chinese AI lab known for efficient, open-weight models. Gained attention for strong performance at lower cost.

### How much does DeepSeek-V4-Pro-0813 cost?

DeepSeek-V4-Pro-0813 costs $0.66 per million input tokens and $1.98 per million output tokens through the DeepSeek API. Cached input is $0.022 per million tokens. Those are the rates for the “Off-peak (all hours outside the published peak windows)” tier; 1 other pricing tier is published for this model. Rates are pay-as-you-go API prices verified against DeepSeek's published pricing on August 18, 2026.

### What benchmark scores did DeepSeek-V4-Pro-0813 get?

DeepSeek-V4-Pro-0813 reports 9 tracked benchmark scores — BullshitBench v2: 35%; SWE-Bench Verified: 80.6%; Terminal-Bench 2.1: 87.9%; Toolathlon-Verified: 74.1%; BrowseComp: 83.4%; CyberGym: 83.3%; Humanity's Last Exam (no tools): 42.7%; Humanity's Last Exam (with tools): 60%; AutomationBench: 31.8%. Tracked scores may come from lab reports or independent benchmarks; source details accompany the benchmark data.

### Is DeepSeek-V4-Pro-0813 open source?

Partly. DeepSeek-V4-Pro-0813 is an open-weight model: the trained weights are free to download and run locally or on your own infrastructure, but the training data and code are not fully released and the license may restrict some commercial uses. It is not open source in the strict sense.

### What came before and after DeepSeek-V4-Pro-0813?

DeepSeek's previous tracked release was DeepSeek-V4-Flash-0731 on Jul 31 2026, 13 days earlier. It was followed by DeepSeek-V4.1-Flash on Sep 10 2026.


---

Canonical page: https://aireleasetracker.com/model/deepseek/deepseek-v4-pro-0813
Site index: https://aireleasetracker.com/llms.txt
Source: AI Release Tracker (https://aireleasetracker.com). Most benchmark scores come from lab launch material; gathered results identify the leaderboard that published them.
