The limited test model runs under the id deepseek-v4.1-flash-expires-on-0910 and is priced like V4 Flash, with a native-multimodal architecture and an official release planned around September 10.
2026-09-09 · by Moe Ameen
In early September 2026, DeepSeek quietly opened a limited-time beta of DeepSeek-V4.1-Flash, posting an "intermediate version" internal test to its official developer communities on September 8. Developers can call it through DeepSeek's existing API — the base URL is unchanged — by selecting the model id `deepseek-v4.1-flash-expires-on-0910`. As the name signals, the test build is scheduled to go offline around September 10, when DeepSeek plans the model's official release (Beijing time). Beta pricing matches [DeepSeek-V4-Flash](/ai-tools/deepseek-v4-flash), with each account limited to about 20 concurrent requests.
The headline change is the architecture. DeepSeek describes V4.1 Flash as a re-architected model that natively integrates multimodal capabilities, so it accepts images alongside text rather than routing vision through a separate experimental build like the earlier [DeepSeek-V4-Flash-Vision-Exp](/ai-tools/deepseek-v4-flash-vision-exp). Alongside multimodality, DeepSeek claims stronger overall performance, faster generation, and lower cost.
The comparison DeepSeek is drawing is the notable part. The company says that across internal and external testing, V4.1 Flash comprehensively surpasses the far larger, more expensive [DeepSeek-V4-Pro](/ai-tools/deepseek-v4-pro-0813) on performance, cost, speed, and total time. DeepSeek also signaled a routing plan: after V4.1 Flash officially launches and before a future V4.1 Pro arrives, it intends to route V4-Pro API requests to V4.1 Flash and bill them at the cheaper V4.1 Flash rate. As this is a short pre-release beta, treat the specifics — exact benchmarks, final pricing, and the release date — as provisional until DeepSeek publishes them in its changelog; confirm against the source before building anything permanent on the test model id, which expires by design.
There are two ways to act on this today, and both point the same direction. First, the practical one: if you already draft on DeepSeek, the V4-Pro-to-V4.1-Flash routing means your scripts, captions, and outlines are about to get cheaper and better on their own. That widens the same bottleneck it always does — you can now afford to write far more than you can produce and publish. Kompozy is the engine that closes that gap. Draft or batch your copy on V4.1 Flash for pennies, bring it into [Kompozy](/) as source, and the engine makes the actual assets a model never will: a [Persona Shorts](/glossary/persona-shorts) avatar video reading your script, a brand-exact [Carousel](/glossary/hyperframes), Quote Graphics, Photo Posts, a Blog Article, and an Email Newsletter, all held to one voice by the [Persona Brief](/glossary/persona-brief), then scheduled across eight social platforms plus blog and email on [autopilot](/glossary/autopilot).
Second, there's a fast play on the news itself. "DeepSeek V4.1 Flash" is a high-intent search this week while developers hunt for what changed and whether the V4 Pro claim holds. A single sharp point of view becomes a week of content inside Kompozy — a short explainer, a carousel breaking down the routing change, captioned clips, and platform-native posts — published while the interest curve is still climbing. DeepSeek is shipping a cheaper, smarter model; Kompozy is how that model's output turns into posts on every platform your audience actually uses. (Kompozy's own copy generation runs on Claude and OpenAI, so V4.1 Flash is your upstream drafting choice and Kompozy is the downstream engine.)
It is a new, re-architected build of DeepSeek's fast Flash tier that natively supports multimodal input (images alongside text). DeepSeek opened a limited beta in early September 2026 under the model id deepseek-v4.1-flash-expires-on-0910 and plans an official release around September 10. Beta pricing matches DeepSeek-V4-Flash.
DeepSeek claims so — it says that across internal and external testing V4.1 Flash comprehensively surpasses the larger V4-Pro on performance, cost, speed, and total time, and plans to route V4-Pro API requests to V4.1 Flash at the cheaper rate after launch. Because it is a short pre-release beta, treat the exact benchmarks as provisional until DeepSeek publishes them.
During the beta it is priced the same as DeepSeek-V4-Flash, with no beta surcharge and roughly 20 concurrent requests per account. Final pricing will be confirmed at the official release; check DeepSeek's changelog for current rates.
No. It is a multimodal text model — it reads images and writes and reasons in text, but it generates no images, video, or audio and publishes nothing. To turn its drafts into finished posts, pair it with a content engine like Kompozy, which generates video, carousels, and images and publishes across the eight social platforms plus blog and email.