DEV Community

Breach Protocol
Breach Protocol

Posted on Originally published at groundtruth.day

DeepSeek is serving a model called V4.1 Flash that it has not announced

A model identified as deepseek-v4.1-flash began answering requests on DeepSeek's API around 8 September 2026, under a model ID that carries its own scheduled expiry date. DeepSeek has published nothing about it: no model card, no technical report, no pricing entry, and no line in its changelog, where the most recent documented release remains V4-Flash-Vision-Exp from 21 August. The company is serving a model it has not announced, and the coverage calling this a launch is running ahead of the evidence.

Key facts

  • What is documented: nothing. DeepSeek's changelog shows no V4.1 entry; the latest is V4-Flash-Vision-Exp, dated 21 August 2026.
  • What is observable: a served model ID with an expiry date attached, plus community-measured throughput.
  • When: first observed 8 September 2026, with the endpoint's own expiry set days later.
  • Primary source: DeepSeek's API changelog.

The expiry date is the most informative detail available, and it points away from a launch rather than toward one. Attaching a scheduled end to a model ID is what a lab does when it wants real traffic against a build for a limited window before deciding what to ship. It is a public test, and the endpoint is meant to vanish.

That has not stopped the story from being reported as a release. It reached the top tier of Hacker News as "DeepSeek launching v4.1 flash cheaper and more capable than v4 pro," a headline asserting two comparative claims — price and capability — that no published document from DeepSeek supports. Neither figure exists anywhere the company controls. A widely repeated statistic putting the new model at a fraction of a competitor's cost for nearly equivalent quality traces back to a third-party leaderboard reading, transmitted through a forum title, about an endpoint with no published pricing.

There is a second, quieter thread worth separating from the first. Users on local-model forums independently noticed that V4 Pro appeared to be receding — described as "soft retired." DeepSeek's changelog does not record any deprecation; V4 Pro's most recent documented update is its general-availability release on 13 August. Two community observations pointing the same direction is a reason to look, not a reason to publish.

Why this is a story rather than an absence of one: DeepSeek has a documented pattern here. Ground Truth has previously covered the company selling access to a checkpoint it had not published, and it is now named in a joint US government advisory on distillation published the same week. A lab that serves models before documenting them is harder to hold to account, because there is no artifact to check a claim against — and the reporting ecosystem fills that vacuum with leaderboard readings and forum titles, which then get cited as though they were specifications.

The analogy is a restaurant serving an unlisted dish. You can order it, and people will tell each other it is excellent and cheap. But there is no menu, no price, and no guarantee it will be there tomorrow — and any review written about it describes something the kitchen never committed to.

For readers deciding what to do with this: if you want to evaluate the model, the endpoint reportedly exists and you can measure it yourself, which is worth more than anyone's summary. If you want to build on it, wait for a model card. The honest position on capability and price today is that nobody outside DeepSeek knows either, and the open-weights status is entirely unestablished — no weights have been published, and whether any will be is not something the company has said. If the weights do not appear, that is the more interesting story, and it would fit the pattern better than a launch does.

There is a broader habit here that goes beyond one company. The gap between when a model starts answering and when it is documented has been widening across the industry, and it inverts how technology reporting is supposed to work. Normally an announcement makes a claim and journalists check it. With undocumented preview endpoints there is no claim to check, so the first descriptions to appear — a forum title, a leaderboard row, a screenshot of a speed test — become the record by default. By the time official documentation arrives, the numbers everyone repeats were set by whoever measured first, under conditions nobody wrote down. The correct response is not to ignore preview endpoints but to label them precisely: a model that is being served is not a model that has been released, and the difference is exactly the part that can be held to account later.


Originally published on Ground Truth, where every claim is checked against the primary source.

Top comments (0)