// AI NEWS · MODEL RELEASE

DeepSeek Ships V4 Pro 0813, Taking Its Flagship Model to General Availability With Frontier-Class Coding at Open-Weight Prices

The 0813 build ends a preview that ran since April, pinning DeepSeek's top tier to a stable version behind the deepseek-v4-pro endpoint — a 1.6-trillion-parameter, MIT-licensed model with a 1M-token context and API rates far below Western frontier labs.

2026-08-12 · by Moe Ameen

What happened

On August 12, 2026, DeepSeek shipped DeepSeek-V4-Pro-0813, the general-availability build of its flagship V4 Pro model. It ends a preview that had run since April 24, 2026, when DeepSeek first released the V4 series with open weights, and it follows the smaller V4-Flash tier, which reached official status on July 31, 2026. The 0813 build now sits behind DeepSeek's `deepseek-v4-pro` API endpoint as a stable, versioned model ID, and its weights remain open under the MIT license.

V4 Pro is a mixture-of-experts model with roughly 1.6 trillion total parameters and about 49 billion active per token. It carries a context window near 1,048,576 tokens and a maximum output around 384,000 tokens, and DeepSeek describes an attention design — a compressed sparse variant plus a heavily compressed one — that the company says cuts single-token inference compute to about 27% and KV cache to roughly 10% of what its V3.2 generation needed at the million-token setting. The API exposes three operating modes: a fast non-thinking mode, a high-reasoning-effort mode, and a max-effort mode.

On the GA build DeepSeek reports strong benchmark figures, weighted toward coding and agentic work: about 80.6% on SWE-bench Verified, 93.5% pass@1 on LiveCodeBench, 90.1% on GPQA Diamond, and a Codeforces rating near 3,206. Pricing carries over from the preview — roughly $0.435 per million input tokens on a cache miss, about $0.0036 per million on a cache hit, and $0.87 per million output — which keeps it well below comparable frontier models from OpenAI, Anthropic, and Google. DeepSeek has not detailed every training change from the preview weights; treat the specific figures as launch-day numbers and confirm them on deepseek.com before relying on them. Like the rest of the family, 0813 is a text-and-reasoning model: it writes, reasons, and codes, but it generates no images, video, or audio and publishes nothing.

Why it matters for creators

  • A frontier-class flagship at open-weight prices resets the drafting budget. When top-tier reasoning costs roughly $0.435/$0.87 per million tokens — with cheap cache hits — high-volume scriptwriting and planning become nearly free for creators.
  • GA plus a stable version ID matters for anyone building a workflow. A pinned model (DeepSeek-V4-Pro-0813) is safe to wire into a recurring content pipeline in a way a moving preview never was.
  • The strengths skew to coding and agentic work, not media. The benchmarks that jumped are SWE-bench, LiveCodeBench, and Codeforces — useful if you build tools around your content, less so if your deliverable is a video or a carousel.
  • MIT-licensed weights keep private, in-house drafting on the table for teams with data-governance concerns about a China-based API — provided they have hardware for a 1.6T model.
  • It changes nothing downstream. 0813 still produces text only: no captions, no reframing, no clips, no brand-voice layer, no scheduler. The last mile of turning drafts into published posts is untouched.

How to act on this with Kompozy

The fastest move on this release is not to open the DeepSeek API — it is to publish your read on it. "The cheapest frontier flagship just hit GA" is a beat your audience is already seeing, so being early with a sharp take beats being thorough a week late. Feed your angle into [Kompozy](/) as a source and it becomes a [Blog Article](/glossary/output-buckets) on what a cheap open-weight flagship changes for content teams, a [Carousel](/glossary/output-buckets) laying out where 0813 sits against closed frontier models, a few captioned shorts, and platform-native posts in your voice through the [Persona Brief](/glossary/persona-brief) — scheduled across the eight social platforms plus blog and email from one queue with [Autopilot](/glossary/autopilot).

The second move is for when you actually draft with it. 0813 is an outstanding upstream brain, and because Kompozy treats every model and generator as an interchangeable input, the fact that top-tier drafting just got cheaper is pure upside — swap it in as your scripting layer without rebuilding anything downstream. Draft a whole week of scripts and outlines in one long 0813 call, then hand each to Kompozy to render the media the model can't: [Persona Shorts](/glossary/persona-shorts) avatar video, brand-exact carousels, [Quote Graphics](/glossary/output-buckets) of the sharpest lines, a formatted blog, and a newsletter — all held to one look by [HyperFrames](/glossary/hyperframes) and published across nine destinations. The model keeps getting cheaper and stronger; the part that still decides whether anyone sees the work is the captioning, formatting, and distribution Kompozy owns. One note: Kompozy's own copy generation runs on Claude and OpenAI, so 0813 is the drafting choice feeding the pipeline, not a swap for it.

Quick takeaways

  • DeepSeek released DeepSeek-V4-Pro-0813, the GA build of its flagship model, on August 12, 2026, ending a preview that ran since April 24, 2026.
  • It is a mixture-of-experts model (~1.6T total / ~49B active parameters) with a ~1M-token context, ~384K-token max output, and MIT-licensed open weights.
  • Reported GA benchmarks skew to coding and agentic work: ~80.6% SWE-bench Verified, 93.5% pass@1 LiveCodeBench, 90.1% GPQA Diamond, ~3,206 Codeforces.
  • API pricing carries over from preview — roughly $0.435/$0.87 per million input/output tokens, with cache hits far cheaper — well below Western frontier models.
  • It is text-only: 0813 drafts scripts and copy, but Kompozy renders the video, carousels, and images and publishes them across nine platforms.

Frequently asked questions

What is DeepSeek V4 Pro 0813, and when was it released?

DeepSeek V4 Pro 0813 is the general-availability build of DeepSeek's flagship V4 Pro model, released on August 12, 2026 after a preview that began April 24, 2026. It is a mixture-of-experts LLM (roughly 1.6 trillion total / 49 billion active parameters) with about a 1-million-token context window and MIT-licensed open weights, served behind DeepSeek's deepseek-v4-pro endpoint.

How much does DeepSeek V4 Pro 0813 cost?

On DeepSeek's first-party API it is roughly $0.435 per million input tokens on a cache miss, about $0.0036 per million on a cache hit, and $0.87 per million output — well below comparable Western frontier models. Because the weights are open under the MIT license, you can also self-host to avoid per-token fees. Confirm current rates on deepseek.com.

Can DeepSeek V4 Pro 0813 make social media content?

It can draft the text — scripts, captions, outlines — but no more. It is a text-and-reasoning model with no image, video, or audio generation and no publishing. Turning its drafts into finished carousels, avatar video, and scheduled multi-platform posts is a separate job handled by a content engine like Kompozy.

How do I act on this release today with Kompozy?

Draft a batch of scripts and outlines in DeepSeek V4 Pro 0813, then bring them into Kompozy to render Persona Shorts, carousels, quote graphics, a blog, and a newsletter in your brand voice, and schedule and publish across the eight social platforms plus blog and email from one queue.

Related news

← All AI news · Get started →