Skip to content
AI EmployeeSuper-AgentsAgent-to-AgentTutorialsPricingBlogStoryContact

DeepSeek V4.1 Flash: Release Date, the Test Model That Expires September 10, and What DeepSeek Has Published

At a glanceQuick answers
Is DeepSeek V4.1 Flash released?
No. As of September 8, 2026 there is no DeepSeek announcement, model card, permanent API id, price or benchmark for V4.1 Flash. What is live is a test id, deepseek-v4.1-flash-expires-on-0910, opened to API users through a community-group notice and named to expire on September 10.
When is DeepSeek V4.1 Flash coming out?
DeepSeek has not said. The test id’s expiry on September 10 is the only date attached to the model, and it is an end date for a test, not a launch date. DeepSeek’s recent releases have arrived 8 to 13 days apart, so a release in September would fit its pattern.
What is different about V4.1 Flash?
According to the relayed notice: a new model structure, native multimodal support, stronger capability, faster speed and lower cost. DeepSeek’s April V4 announcement described the V4 structure as token-wise compression plus DeepSeek Sparse Attention with a 1M context; what changed in 4.1 has not been published.
What is DeepSeek's current lineup?
Three API models as of September 8: deepseek-v4-flash (serving DeepSeek-V4-Flash-0731, released July 31), deepseek-v4-pro (serving DeepSeek-V4-Pro-0813, released August 13) and the experimental deepseek-v4-flash-vision-exp (released August 21). V4-Flash is 284B total and 13B active parameters with a 1M context, per DeepSeek’s April announcement.
Hand-drawn sketch of two solid boxes labeled V4 FLASH with a JUL 31 tag and VISION-EXP with an AUG 21 tag, an arrow to a dashed box labeled V4.1 FLASH with an eye symbol and NEW STRUCTURE, and an hourglass with amber sand labeled EXPIRES 09/10
Fig 0Two models shipped with dates on the tag. The third is a dashed line with an hourglass over it, because the only official thing about it is its expiry.

DeepSeek V4.1 Flash has a date attached to it, and the date is an expiry. On September 8, 2026, at about 3 pm Beijing time, DeepSeek posted a notice in its official user community group: an “intermediate version” of V4.1 Flash was open for testing under the API model id deepseek-v4.1-flash-expires-on-0910, billed at the same rate as deepseek-v4-flash, capped at 20 concurrent requests per account. The notice, as relayed by Chinese technology press within the hour, promised a new model structure, native multimodal support, more capability, more speed and lower cost. It did not promise a release date, and DeepSeek has published nothing since.

We checked the channels DeepSeek actually uses. As of September 8, 2026, the API news page ends at the August 21 release of V4-Flash-Vision-Exp. The models and pricing page lists three models, none of them 4.1. The change log has no 4.1 entry. The deepseek-ai organization on Hugging Face has no 4.1 repository, and the newest model there is still V4-Flash-Vision-Exp. The company’s X account has said nothing. DeepSeek’s own April announcement carries a line worth keeping in mind here: “please rely only on our official accounts for DeepSeek news.”

So this page does what our Qwen 4, Gemini 4 and Grok 4.7 trackers do: every claim dated, the company’s documents separated from the relayed notice and both separated from community tests, updated the same day the story moves. DeepSeek is a different kind of tracker than Grok. There is no founder posting dates, no leaked benchmark sheet, no Reddit war. There is a model id with a two-day life and a release history that is unusually regular. That history is most of what we can say.

On this page · 8 sectionsOpen
  1. What we actually know
  2. What the notice said, and who carried it
  3. The rumor chain, dated
  4. DeepSeek’s cadence
  5. What 4.1 would land on top of
  6. What we are watching for
  7. The tracker
  8. Update log
Key points7 · 14 min full read
  1. DeepSeek V4.1 Flash is not released as of September 8, 2026. DeepSeek’s API news page, its model list, its change log, its Hugging Face organization and its X account contain no announcement, model card, price or spec for a 4.1 model.
  2. What exists is a test. On September 8 at about 3 pm Beijing time, DeepSeek posted a notice in its official community group opening an ‘intermediate version’ of V4.1 Flash under the API model id deepseek-v4.1-flash-expires-on-0910, billed at V4-Flash rates, limited to 20 concurrent requests per account. The id says it expires September 10.
  3. The notice, as relayed by IT之家, Phoenix Tech and Machine Heart, describes a new model structure, native multimodal support, more capability, more speed and lower cost. None of that is in a DeepSeek document, and no parameter count, context window or benchmark has been published.
  4. The most telling detail is a question in DeepSeek’s feedback form: whether the V4.1 Flash intermediate version can fully replace the online DeepSeek V4 Pro. A Flash-class model replacing Pro would be the story if it happens.
  5. Speed claims of about 400 tokens per second and 3.9 to 6 times faster than V4-Flash-Vision-Exp come from one X user’s self-tests, relayed by press. They are unverified.
  6. DeepSeek’s own cadence says a release is plausible soon: V4 Preview April 24, V4-Flash-0731 on July 31, V4-Pro-0813 on August 13, V4-Flash-Vision-Exp on August 21. The last three landed 13 and 8 days apart.
  7. This is a living tracker. When DeepSeek publishes an announcement, a model card or a price, the confirmed facts land here the same day next to a scorecard of every claim above.

§ 01What we actually know

Claim Status
A V4.1 Flash test model is live on the API Confirmed: id deepseek-v4.1-flash-expires-on-0910, per DeepSeek’s community-group notice as relayed by IT之家, Phoenix Tech and Machine Heart
Billing equals deepseek-v4-flash Per the notice, relayed; DeepSeek’s pricing page does not list the id
20 concurrent requests per account Per the notice, relayed
The test ends September 10 The model id says so; DeepSeek has not published a date
New model structure Per the notice; nothing published on what changed
Native multimodal support Per the notice; V4-Flash-Vision-Exp already accepts images via an added encoder
Stronger, faster, cheaper Per the notice; no benchmark, throughput or price published
About 400 tokens per second Community self-test relayed by Machine Heart; unverified
3.9 to 6 times faster than V4-Flash-Vision-Exp One X user’s self-tests (@NFT_Chen) relayed by Machine Heart; unverified
“Can it replace the online V4 Pro?” A question in DeepSeek’s own tester feedback form, per IT之家
Release date Nothing. DeepSeek has not announced one
Parameters, context window, permanent API id, model card, price Nothing exists
DeepSeek’s current lineup deepseek-v4-flash (V4-Flash-0731), deepseek-v4-pro (V4-Pro-0813), deepseek-v4-flash-vision-exp, per the pricing page
Table 1DeepSeek V4.1 Flash claims vs verifiable status, September 8, 2026

§ 02What the notice said, and who carried it

DeepSeek did not publish the notice on a public page, so the primary record is the press that quoted it. IT之家 posted at 16:05 Beijing time, citing readers who saw the message in DeepSeek’s official group: an intermediate version of DeepSeek V4.1 Flash is in internal testing, it uses a new model structure with native multimodal support, it is stronger, faster and cheaper, testers keep base_url unchanged and set the model to deepseek-v4.1-flash-expires-on-0910, billing matches deepseek-v4-flash, and each account is limited to 20 concurrent requests. IT之家 also reported the feedback form’s question about replacing V4 Pro. Phoenix Tech carried the same points at 16:17. Machine Heart, at 18:00, put the group message at about 3 pm and added that “internal test” was generous: any API user could add the model name and call it.

Three things about the notice are worth reading carefully. First, “intermediate version” is DeepSeek’s own framing. This is not the release model; it is a checkpoint offered for feedback, and the expiry in the id enforces that. Second, “new model structure” is a stronger claim than DeepSeek made for V4-Flash-0731, whose model card describes it as the official release of V4-Flash with enhanced agentic capabilities and the same structure as the DSpark variant. A structural change between 4 and 4.1 is a bigger step than the version number suggests. Third, “native multimodal” is a distinction from the current approach. V4-Flash-Vision-Exp added a vision encoder and aligner to the text model and kept training. Native means the images are part of the base design. If true, it explains why DeepSeek would call it a new structure.

§ 03The rumor chain, dated

The chain is short, because DeepSeek does not leak the way Western labs do. What it does instead is ship with the date in the file name.

Date Event Source type
Apr 24 DeepSeek-V4 Preview: V4-Pro 1.6T/49B active, V4-Flash 284B/13B active, 1M context, open weights DeepSeek, official
Jul 31 DeepSeek-V4-Flash-0731 official release, superseding the preview DeepSeek, official
Aug 13 DeepSeek-V4-Pro GA (V4-Pro-0813) with DeepSeek Harness; peak and off-peak API pricing announced DeepSeek, official
Aug 16 New pricing takes effect, 16:00 UTC DeepSeek, official
Aug 21 DeepSeek-V4-Flash-Vision-Exp live on the API; Harness 0.1.1 DeepSeek, official
Aug 31 V4-Flash-Vision-Exp open weights on Hugging Face, MIT license DeepSeek, official
Sep 1 to 7 DeepSeek Harness 0.1.3 pre-release builds tagged on GitHub DeepSeek, official (pre-release)
Sep 8, 3 pm CST Community-group notice: V4.1 Flash intermediate version, expires-on-0910 id DeepSeek notice, relayed by press
Sep 8, 16:05 CST IT之家 reports the notice and the V4 Pro replacement question Press
Sep 8, 18:00 CST Machine Heart reports community speed tests, 400 tokens per second Press, citing users
Sep 8 DeepSeek news page, pricing page, change log, Hugging Face, X: no “4.1” Our own read
Sep 10 Test id expires, per its name Implied by DeepSeek’s id
Table 2The DeepSeek V4-era chain, dated
One test id in a year of dated releasesTimeline from the April 24 V4 Preview through the July 31 V4-Flash-0731 release, the August 13 V4-Pro GA, the August 21 Vision-Exp API launch, the August 31 open weights, the September 8 V4.1 Flash test notice highlighted, and the September 10 expiry of the test idApr 24V4 Preview, 1M contextJul 31V4-Flash-0731 releaseAug 13V4-Pro GA and HarnessAug 21Vision-Exp on the APIAug 31Vision-Exp open weightsSep 8V4.1 Flash test id appearsSep 10Test id expiresOne test id in a year of dated releasesTimeline from the April 24 V4 Preview through the July 31 V4-Flash-0731 release, the August 13 V4-Pro GA, the August 21 Vision-Exp API launch, the August 31 open weights, the September 8 V4.1 Flash test notice highlighted, and the September 10 expiry of the test idApr 24V4 Preview, 1M contextJul 31V4-Flash-0731 releaseAug 13V4-Pro GA and HarnessAug 21Vision-Exp on the APIAug 31Vision-Exp open weightsSep 8V4.1 Flash test id appearsSep 10Test id expires
Fig 1One test id in a year of dated releases

The expiry is the only date, and it is not a launch date. expires-on-0910 tells testers when the checkpoint disappears. It does not tell anyone when a release model arrives, and DeepSeek’s dated suffixes have always described the release itself (0731, 0813), never a countdown. Reading September 10 as a release date is the mistake this page exists to prevent.

The speed numbers are one person’s tests. Machine Heart credits the 400 tokens per second figure to unnamed community members and the 3.9 to 6 times speedups to X user @NFT_Chen, comparing the test model against V4-Flash-Vision-Exp on a 49K-token retrieval task (5.2 times) and SVG code generation (6 times). Those are plausible for a model DeepSeek itself calls faster, and they are self-tests of an intermediate checkpoint over a two-day window. We list them as reported. The Gemini 3.8 Flash comparison Machine Heart draws is the comparison DeepSeek’s launch table will eventually make or avoid; until then it is a tweet.

The V4 Pro question is the real signal. DeepSeek asked its testers whether this Flash-class model can fully replace V4 Pro online. V4-Pro is the 1.6T total, 49B active flagship; V4-Flash is 284B total, 13B active. If DeepSeek is seriously weighing a Flash replacing Pro in its own service, the new structure is doing a great deal of work, and it would fit a company whose April announcement led with “cost-effective 1M context” rather than raw scale.

§ 04DeepSeek’s cadence

The reason a September release is plausible has nothing to do with the notice. It is the calendar.

Model Announced by DeepSeek Days since previous
DeepSeek-V3 December 26, 2024 n/a
DeepSeek-R1 January 20, 2025 25
DeepSeek-V3-0324 March 25, 2025 64
DeepSeek-R1-0528 May 28, 2025 64
DeepSeek-V3.1 August 21, 2025 85
DeepSeek-V3.1-Terminus September 22, 2025 32
DeepSeek-V3.2-Exp September 29, 2025 7
DeepSeek-V3.2 December 1, 2025 63
DeepSeek-V4 Preview April 24, 2026 144
DeepSeek-V4-Flash-0731 July 31, 2026 98
DeepSeek-V4-Pro-0813 August 13, 2026 13
DeepSeek-V4-Flash-Vision-Exp August 21, 2026 8
DeepSeek V4.1 Flash Test id September 8; no release 18 and counting
Table 3DeepSeek releases since V3 and the gap since the previous release
Days between DeepSeek releases, V3 to todayBar chart of days between consecutive DeepSeek releases from V3 in December 2024 to the V4-Flash-Vision-Exp release in August 2026, with the open gap to the September 8 V4.1 Flash test highlightedR125V3-032464R1-052864V3.185V3.1-Terminus32V3.2-Exp7V3.263V4 Preview144V4-Flash-073198V4-Pro-081313Vision-Exp8V4.1 test (open, Sep 8)18Days between DeepSeek releases, V3 to todayBar chart of days between consecutive DeepSeek releases from V3 in December 2024 to the V4-Flash-Vision-Exp release in August 2026, with the open gap to the September 8 V4.1 Flash test highlightedR125V3-032464R1-052864V3.185V3.1-Terminus32V3.2-Exp7V3.263V4 Preview144V4-Flash-073198V4-Pro-081313Vision-Exp8V4.1 test (open, Sep 8)18
Fig 2Days between DeepSeek releases, V3 to today

Dates come from the DeepSeek API news index and change log; V4-Flash-0731 is dated by its July 31 change log entry, and V3-0324 by its news page rather than the March 24 in its name. Two readings of the table. The long gaps are between generations: 144 days from V3.2 to V4 Preview, 98 from the preview to the first final V4 model. The short gaps are inside a generation: 13 days from V4-Flash-0731 to V4-Pro, 8 from V4-Pro to Vision-Exp. A 4.1 that is a structural change sits somewhere between those two rhythms. The September 8 test is 18 days after Vision-Exp; if V4.1 Flash ships in September, it will be the fastest structural update DeepSeek has made, and the expiring test id is the first time DeepSeek has previewed a model this way rather than shipping it.

§ 05What 4.1 would land on top of

The V4 lineup is well documented, which is what makes the 4.1 silence stand out. DeepSeek’s April 24 announcement gave the architecture: V4-Pro at 1.6T total and 49B active parameters, V4-Flash at 284B total and 13B active, both with a 1M context as the default across all official DeepSeek services, built on token-wise compression plus DeepSeek Sparse Attention. The July 31 V4-Flash-0731 card calls itself the official release of V4-Flash and reports it beating the V4-Pro preview on agent benchmarks despite the far smaller active count (Terminal Bench 2.1 at 82.7, DeepSWE at 54.4, Toolathlon-Verified at 70.3 on DeepSeek’s own table). The card’s weight total reads about 304B because the release ships with a speculative decoding module attached; the announced base is 284B. The August 13 V4-Pro GA post added reasoning effort levels for both models, native OpenAI Responses API support, and the pricing change: peak and off-peak rates, off-peak 50 percent lower, effective August 16 at 16:00 UTC. Current figures are on the pricing page; the test id is billed at whatever deepseek-v4-flash costs at the hour you call it.

The agent tooling is moving too. DeepSeek Harness 0.1.1 shipped with Vision-Exp on August 21, and the GitHub repository tagged 0.1.3 pre-release builds every day from September 1 to 7. A new model with a new structure arriving alongside a harness release would match how V4-Pro landed. Our record of the wider open-weight wave this lands in is in the Qwen3.8-Flash-Next and GLM-5.3-Flash posts.

§ 06What we are watching for

We run production work on three model families and read the open-weight labs closely, so this list is an operator’s.

  • An entry on the API news page. Every DeepSeek release since V3 has one. The moment a 4.1 page appears at api-docs.deepseek.com/news, this stops being a rumor page.
  • A permanent model id. deepseek-v4.1-flash without an expiry, or a dated release suffix like the ones on 0731 and 0813, on the models and pricing page.
  • Weights and a card on Hugging Face. V4-Flash-Vision-Exp’s weights followed its API release by ten days. A card would give the first parameter count and confirm or retire the “new structure” line.
  • September 10. If the test id expires and nothing replaces it, that is a scorecard row. If a second test id appears with a later expiry, that is a pattern.
  • A price. “Lower cost” is in the notice. DeepSeek’s V4 pricing change was announced eight days before it took effect; a 4.1 price would likely arrive the same way.
  • The V4 Pro decision. Whether DeepSeek’s release post says anything about the online service’s default model, and whether deepseek-v4-pro keeps pointing at 0813.
  • Multimodal spec. Whether “native” means the release model accepts images without a separate vision-exp id, which would retire the August 21 experiment.

§ 07The tracker

This page updates when facts change, not when posts get louder. As of September 8, 2026: DeepSeek V4.1 Flash exists as a test model id with a two-day life, described in a group notice, tested by strangers and announced nowhere. No release date, no spec, no benchmark and no price exist in any DeepSeek document. The likeliest picture, on DeepSeek’s own rhythm, is a release in September with the date in its name. When DeepSeek moves, the confirmed facts and a graded scorecard of every claim above land here the same day.

§ 08Update log

This is a living page; when the story moves, the update lands here.

As of September 8, 2026: page opened. Test id deepseek-v4.1-flash-expires-on-0910 live per DeepSeek’s community notice; no official announcement to log.

Frequently asked6 questions

Q1Has DeepSeek announced V4.1 Flash?

Not on any channel it uses for announcements. As of September 8, 2026 the last entry on the DeepSeek API news page is the August 21 release of DeepSeek-V4-Flash-Vision-Exp, the pricing page lists three models with no 4.1, the change log has no 4.1 entry, the deepseek-ai organization on Hugging Face has no 4.1 repository, and the company’s X account has posted nothing about it. The only DeepSeek-authored text about V4.1 Flash is a notice in its official user community group on September 8, which was relayed by Chinese technology press within an hour.

Q2What does deepseek-v4.1-flash-expires-on-0910 mean?

It is a temporary API model id. DeepSeek’s notice told testers to keep their base_url and set the model name to deepseek-v4.1-flash-expires-on-0910. Billing is the same as deepseek-v4-flash and each account is limited to 20 concurrent requests. The suffix is the expiry: the id stops working on September 10. DeepSeek has used dated suffixes before for release versions, such as V4-Flash-0731 and V4-Pro-0813, but an ‘expires-on’ suffix is new and it marks a test window, not a release.

Q3Is V4.1 Flash faster than V4-Flash?

Reported, not confirmed. Machine Heart’s September 8 write-up cites community users reaching about 400 tokens per second and one X user, @NFT_Chen, measuring 3.9 to 6 times faster end-to-end than V4-Flash-Vision-Exp, including 5.2 times faster on a 49K-token retrieval task and 6 times faster on SVG generation. Those are self-tests on a test model during a two-day window. DeepSeek has published no throughput figure. If the release model ships with a spec sheet, the speed row goes into the tracker with its source.

Q4Will V4.1 Flash replace V4 Pro?

DeepSeek is asking that question itself. Its tester feedback form, per IT之家’s report, asks whether the V4.1 Flash intermediate version can fully replace the online DeepSeek V4 Pro. That would be unusual: V4-Pro is DeepSeek’s 1.6T total, 49B active flagship and V4-Flash is the 284B total, 13B active efficient model. A Flash-class model good enough to retire Pro from the online service would say more about the new structure than any benchmark. Nothing has been decided that DeepSeek has published.

Q5What is DeepSeek's release pattern?

Fast and dated, with the dates in the model names. In 2026 alone: V4 Preview on April 24, V4-Flash-0731 on July 31, V4-Pro-0813 with the DeepSeek Harness on August 13, V4-Flash-Vision-Exp on the API on August 21 and its open weights on August 31. The three summer releases landed 13 and 8 days apart. On that rhythm a September V4.1 release is plausible; on DeepSeek’s history, a test model with an expiry date has not always turned into a release of the same name.

Q6Should I wait for DeepSeek V4.1 Flash?

If you use DeepSeek’s API, the test id is available now at V4-Flash pricing and vanishes on September 10, so trying it costs nothing beyond tokens. If you are choosing where your agents run, the model underneath matters less than the layer around it: memory that survives the conversation, a task board, an inbox, approvals a human can see. CellCog does not route to DeepSeek today; our tiers run on Gemini 3.8 Flash, Claude Fable 5.1 and Opus 5, and a standing AI employee keeps its role across every engine swap. Plans start at $8 a month; you pay for the work, not the hire, and the cost depends purely on how much work you assign.

Published 08 September 2026 All Choosing a platform →