• Latest
  • Trending
  • All
Answer card: DeepSeek replaced the weights behind deepseek-v4-pro with DeepSeek-V4-Pro-0813 on 12 August 2026, keeping the same endpoint and the same price of 0.435 dollars per million input tokens and 0.87 per million output, with no change log entry and no 0813 weights on Hugging Face.

DeepSeek V4 Pro 0813 is live, at 3x the Flash price

12 August 2026
Answer card stating that Mullvad announced on 3 September 2026 that it is shutting down its public encrypted domain name system servers on 2 November 2026 and sponsoring the Quad9 Foundation instead, with 194.242.2.2 and its five sibling addresses all going away, and virtual private network customers unaffected.

Mullvad’s DNS servers go dark on 2 November, and Quad9 blocks no ads

5 September 2026
OpenAI announcement image for GPT-6 Astra, a spiral galaxy of white, blue and amber points of light curling around a bright core on a near black star field.

GPT-6 Astra lists at $10 and $50, 2.5x what GPT-5.6 Sol costs

6 September 2026
Google's official announcement image for the release, reading Introducing Gemini 3.8 Flash and 3.8 Flash Cyber in black type over a pale blue background with a blurred white chevron and the four colour Gemini spark below.

Gemini 3.8 Flash keeps the price and the 1 January cliff

3 September 2026
Answer card stating that Anthropic announced Enterprise Frontier Safeguards on 1 September 2026, that activity data used for misuse monitoring moves into cloud storage the customer controls under the customer own encryption keys, that Anthropic charges nothing for the feature while the cloud provider bills storage and egress, and that the phased rollout starts later in autumn 2026 with interim zero data retention on Fable 5 and Fable 5.1 for eligible customers.

Anthropic moves retention into your own cloud, for 30 days

3 September 2026
Official Google diagram of a client connection in three numbered steps: a DNS lookup with a query and an address, a TLS ClientHello and ServerHello, then a content exchange with a website. A callout on the DNS step reads 25% of global web traffic is now protected by encrypted DNS, and a callout beside an Android phone on the ClientHello step reads Android 17 supports ECH GREASE by default.

Android 17 hides the SNI, not your DNS or destination

3 September 2026
Still frame from the Claude Fable 5.1 launch video showing model-designed protein binders in orange docked against twelve grey target proteins, rendered as ESMFold2 structure predictions.

Claude Fable 5.1 breaks forced tool use, cuts cache 75%

1 September 2026
Answer card stating that on 31 August 2026 the European Commission designated ChatGPT a Very Large Online Search Engine under the Digital Services Act, the first conversational AI service classified that way, because it answers user prompts and queries including by searching the web, with OpenAI having declared roughly 159.1 million average monthly users in the European Union for ChatGPT search.

The EU now calls ChatGPT a very large search engine

3 September 2026
Answer card stating that on 31 August 2026 the Department of War added OpenAI ChatGPT Mil and Starshield AI Grok for Government to the GenAI.mil portal alongside Google Gemini, all three accredited at Impact Level 5 for Controlled Unclassified Information, with 1.7 million unique users onboarded out of roughly 3 million eligible personnel, and ChatGPT Mil currently serving GPT-5.4 Terra with GPT-5.6 Terra said to be rolling out.

ChatGPT Mil and Grok reached IL5 on GenAI.mil

3 September 2026
Answer card stating that Anthropic opened a research preview of the Model Hardware Standard on 27 August 2026, standardising the driver layer between an operating system and a laboratory instrument with read and write primitives plus discovery and safety limits, reachable through MCP as well as a command line and code files, with no public specification published.

Anthropic’s Model Hardware Standard is gated, and sits under MCP

3 September 2026
Official Cohere key art for the Parse 5 launch: the Cohere mark and the wordmark Parse with a superscript 5 in white, centred on a soft out of focus gradient of deep blue, violet and amber curves.

Cohere Parse 5 is $1.50 per 1,000 pages, on three of five dimensions

3 September 2026
Title card from the OpenAI announcement video: a man sits on a blue sofa in a loft with tall windows and potted plants, a laptop open on the coffee table in front of him, with the words WebMCP in ChatGPT in large white type across the lower left.

WebMCP in ChatGPT needs GPT-5.6 Sol or Terra

3 September 2026
The Agentic Coding section of the official Hy4 preview benchmark appendix published by Tencent, a table comparing Hy3 and Hy4 preview against DeepSeek V4 Pro 0813, Qwen 3.8 Max, GLM 5.3, Kimi K3, GPT 5.6 Sol and Claude Opus 5 across SWE-bench Multilingual, SWE-bench Pro, DeepSWE, three SWE Atlas tasks, SWE-Marathon, Terminal-Bench 2.1, NL2Repo-Bench, CyberGym, ProgramBench, PostTrainBench and Harbor-Index.

Tencent’s 770B Hy4 tops one benchmark row in 46

3 September 2026
  • About
  • Contact
  • Privacy
  • Legal
Sunday, September 6, 2026
  • Login
Packet Nebula
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About
No Result
View All Result
Packet Nebula
No Result
View All Result
Home Dev

DeepSeek V4 Pro 0813 is live, at 3x the Flash price

by stephane
12 August 2026
in Dev
0
Answer card: DeepSeek replaced the weights behind deepseek-v4-pro with DeepSeek-V4-Pro-0813 on 12 August 2026, keeping the same endpoint and the same price of 0.435 dollars per million input tokens and 0.87 per million output, with no change log entry and no 0813 weights on Hugging Face.
491
SHARES
1.4k
VIEWS
Share on FacebookShare on Twitter

Nothing in your code has to change, again. On 12 August DeepSeek pointed deepseek-v4-pro at a new build called DeepSeek-V4-Pro-0813, same endpoint, same 0.435 dollars in and 0.87 out, new weights underneath. We found out the way everyone else did, by noticing a table on the pricing page had changed, because there's no change log entry and no blog post. Flash got the same treatment on 31 July, and at least that one came with a note. The swap itself is routine by now. What's worth your attention is what it does to the choice between the two DeepSeek models, because for the past two weeks the cheap one had been beating the expensive one, and that's over.

The short answer

DeepSeek-V4-Pro-0813 now answers calls to deepseek-v4-pro. The endpoint, the price and the 1M context are unchanged, and the swap arrived with no change log entry. The 0813 weights are not on Hugging Face, which breaks the pattern every previous V4 release set. Pro is faster than Flash on DeepSeek’s own agent benchmarks again, at a little over three times the cost per token.

12 Augweights swapped in place
3.1xPro token price against Flash
500concurrent requests, Flash gets 2500
Answer card: DeepSeek replaced the weights behind deepseek-v4-pro with DeepSeek-V4-Pro-0813 on 12 August 2026, at the same endpoint and the same price of 0.435 dollars per million input and 0.87 per million output, with no change log entry and no 0813 weights on Hugging Face.
A version string in a docs table is the whole announcement so far.

What actually changed

We compared the models and pricing page against its own archived copy. On 9 August the model version column read DeepSeek-V4-Pro. Today it reads DeepSeek-V4-Pro-0813. That’s it. That’s the announcement.

One other row moved with it, and it’s the one developers were waiting on. The Responses API was flash-only until now, with a footnote promising deepseek-v4-pro support in early August. That footnote is gone and the row is a tick. So if you were holding a Responses API migration because your hard calls go to Pro, you can stop holding.

Everything else on the page held still. Same 1M context, same 384K ceiling on output, same prices to four decimal places, same concurrency limits. The change log still ends at 31 July, which was the Flash update. Nothing about 0813.

Checklist comparing what changed on the DeepSeek pricing page between 9 and 12 August 2026, namely the model version string and Responses API support for deepseek-v4-pro, against what did not change, namely the prices, the 500 request concurrency limit, the absence of a change log entry and the absence of 0813 weights on Hugging Face.
Two ticks moved. The right-hand column is the part worth arguing about.

The weights broke the pattern

Here’s the bit we didn’t expect. Every V4 release so far shipped with weights. The April preview put both Pro and Flash on Hugging Face under MIT on day one. Flash 0731 did the same, ungated, with a vLLM recipe in the model card.

There is no DeepSeek-V4-Pro-0813 repository. We checked the Hugging Face API directly rather than trusting a search box, and the newest thing on the deepseek-ai account is still Flash 0731, uploaded 1 August. So the weights you can download today are the April preview. Not what the API is serving.

Maybe they land next week and this paragraph ages badly. I’d guess they do, honestly, because opening the weights is most of DeepSeek’s leverage and they’ve never skipped it. But if you self-host V4-Pro and you’ve been telling people it matches the hosted model, that stopped being true on 12 August, and nobody sent you a note.

Pro takes the lead back from Flash

Two weeks ago the awkward fact about DeepSeek’s lineup was that the small model beat the big one. Flash 0731 outscored the V4-Pro preview on every agent benchmark DeepSeek published, at a third of the price, which made Pro difficult to recommend for anything.

The numbers going round for 0813 fix that. Terminal Bench 2.1 at 87.9 where the preview managed 72.1. DeepSWE at 62.7 from 12.8. AutomationBench 31.8 from 12.8, DSBench-Hard 67.2 from 31.1. Those come from a DeepSeek chart circulating on X and reported by trade press, and we could not find them on any DeepSeek page. Read them as DeepSeek’s claim about DeepSeek, because that’s what they are.

Bar chart of Terminal Bench 2.1 scores reported by DeepSeek for three of its own builds: V4-Pro preview at 72.1 in April, V4-Flash-0731 at 82.7 in July, and V4-Pro-0813 at 87.9 in August.
Ahead again. By 5.2 points, for a little over three times the price per token.

Put the price next to it and the decision gets simpler, not harder. Pro is $0.435 per million input and $0.87 output. Flash is $0.14 and $0.28. That ratio is exactly 3.107 on both sides, which is a suspiciously tidy number and probably deliberate. You’re paying triple for about five points on a benchmark DeepSeek ran itself.

The concurrency limit is the part people miss. Pro allows 500 concurrent requests, Flash allows 2500. If you’re running a batch job or a fan-out agent, that ceiling will hurt you long before the bill does.

The price warning is still sitting there

Same page, footnote one, and it hasn’t moved since early August:

We plan to raise the overall pricing for DeepSeek API services in the near future, with a significant increase expected. Please plan your usage accordingly. The specific pricing plan will be subject to official notice.

Worth knowing what that footnote replaced. Earlier in August the same slot held a concrete plan: peak and off-peak billing, 2x the regular rate between 09:00 and 12:00 and again 14:00 to 18:00 Beijing time. That’s gone, swapped for a blanket warning with no number and no date. Whether that’s better or worse depends entirely on how much of your traffic was going to land in those windows.

Either way, don’t model a year of spend on today’s rates. DeepSeek has told you in writing that they’re temporary.

Should you switch anything

If your Pro calls are already going to deepseek-v4-pro, you’ve switched. That’s the nature of an in-place swap, and it’s the same trap we wrote about when the legacy aliases retired in July. Re-run whatever evals you trust. The results you have on file describe the April preview.

If you’d moved hard work to Flash because Pro looked bad value, it’s worth another look, though not automatically. Five benchmark points is real on tasks that were failing outright and invisible on tasks that were already passing. Try it on the calls that actually hurt.

And if you self-host, sit tight. The build worth downloading isn’t downloadable yet.

Sources

Model version, pricing, concurrency limits and the price increase notice are from DeepSeek’s own models and pricing page, compared against the archived copy of 9 August 2026 for the diff. The GA description, the endpoint creation time of 15:42 UTC on 12 August and the context and output limits are from the OpenRouter listing and its public models API. Weight availability was checked against the deepseek-ai account on Hugging Face. The 0813 benchmark figures are DeepSeek-reported, circulated via a widely shared post on X and covered by Wccftech, and we have not reproduced them. The preview and Flash 0731 scores are from DeepSeek’s published table of 31 July.

Frequently asked questions

What is DeepSeek-V4-Pro-0813?

It is the build now sitting behind the deepseek-v4-pro model name in DeepSeek's API. The models and pricing page listed the version as plain DeepSeek-V4-Pro as recently as 9 August and shows DeepSeek-V4-Pro-0813 on 12 August. OpenRouter listed a matching deepseek-v4-pro-0813 endpoint at 15:42 UTC on 12 August and describes it as the GA release of DeepSeek V4 Pro, which is the closest thing to an announcement anyone has.

Do I need to change my API calls?

No. The model name is still deepseek-v4-pro, the base URLs are unchanged, and the 1M context and 384K maximum output are the same. That is exactly why it is worth a note in your own log: anything already pointed at that name picked up new weights without an opt-in. If you keep an eval suite from before 12 August, it describes a model you are no longer talking to.

How much does DeepSeek V4 Pro cost now?

Per million tokens, DeepSeek publishes $0.003625 on a cache hit, $0.435 on a cache miss and $0.87 for output. Those numbers did not move with the 0813 swap. They are 3.1x the deepseek-v4-flash rates of $0.14 and $0.28. The same page carries a warning that DeepSeek plans to raise overall API pricing in the near future, with a significant increase expected, so treat today's figures as current rather than settled.

Are the DeepSeek-V4-Pro-0813 weights open?

Not as of 12 August. There is no DeepSeek-V4-Pro-0813 repository on Hugging Face, and the most recent upload on the deepseek-ai account is still DeepSeek-V4-Flash-0731 from 1 August. The April V4-Pro preview weights are up under the MIT licence, so if you self-host today you are running the preview, not the build the API now serves.

Should I use V4 Pro or V4 Flash?

Flash for almost everything, Pro when the task is genuinely hard. On DeepSeek's own Terminal Bench 2.1 numbers the Pro 0813 build scores 87.9 against 82.7 for Flash 0731, so the flagship is ahead again after two weeks of being behind. That is 5.2 points for 3.1x the token price and a fifth of the concurrency, since Pro is capped at 500 concurrent requests against 2500 for Flash.

Tags: aiapideepseekdev-toolsllmnewsopen-weights
Share196Tweet123
stephane

stephane

  • Trending
  • Comments
  • Latest
Answer card: Proton Lumo 2.0 is private by policy, not by locality. Saved history is locked so even Proton cannot read it, but the prompt is decrypted on a Proton EU server to answer it, then forgotten.

Proton Lumo 2.0 review: how private is it, really?

3 September 2026
Google's official announcement image for the release, reading Introducing Gemini 3.8 Flash and 3.8 Flash Cyber in black type over a pale blue background with a blurred white chevron and the four colour Gemini spark below.

Gemini 3.8 Flash keeps the price and the 1 January cliff

3 September 2026
Answer card stating that Anthropic announced Enterprise Frontier Safeguards on 1 September 2026, that activity data used for misuse monitoring moves into cloud storage the customer controls under the customer own encryption keys, that Anthropic charges nothing for the feature while the cloud provider bills storage and egress, and that the phased rollout starts later in autumn 2026 with interim zero data retention on Fable 5 and Fable 5.1 for eligible customers.

Anthropic moves retention into your own cloud, for 30 days

3 September 2026
Answer card: JWTs are not encrypted, anyone can read them; the signature proves who issued the token, not who may read it.

Are JWTs encrypted? No, and the difference will bite you

0
Answer card: a random 8 character password falls in under 2 hours offline, while 16 random characters hold for 1.4 trillion years at the same speed.

How long does it take to crack a password in 2026?

0
Answer card: three DNS records decide if your mail lands or bounces; SPF lists allowed senders, DKIM signs messages, DMARC sets the failure policy.

SPF, DKIM and DMARC explained: the records your email needs

0
Answer card stating that Mullvad announced on 3 September 2026 that it is shutting down its public encrypted domain name system servers on 2 November 2026 and sponsoring the Quad9 Foundation instead, with 194.242.2.2 and its five sibling addresses all going away, and virtual private network customers unaffected.

Mullvad’s DNS servers go dark on 2 November, and Quad9 blocks no ads

5 September 2026
OpenAI announcement image for GPT-6 Astra, a spiral galaxy of white, blue and amber points of light curling around a bright core on a near black star field.

GPT-6 Astra lists at $10 and $50, 2.5x what GPT-5.6 Sol costs

6 September 2026
Google's official announcement image for the release, reading Introducing Gemini 3.8 Flash and 3.8 Flash Cyber in black type over a pale blue background with a blurred white chevron and the four colour Gemini spark below.

Gemini 3.8 Flash keeps the price and the 1 January cliff

3 September 2026
  • About
  • Contact
  • Privacy
  • Legal

Copyright © 2026 Stephane Cardon.

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About

Copyright © 2026 Stephane Cardon.