← Founder Notes
Archive

The frontier open model race just moved to sparse math. deepseek v4.1 flash, reported september 14,…

Yethikrishna ROriginal on Threads

the frontier open model race just moved to sparse math. deepseek v4.1 flash, reported september 14, runs a 552-billion-parameter backbone with only 8 billion active per token, under an mit license, and beats deepseek's own v4 pro on agentic coding benchmarks.

sparsity became the moat.

Context

DeepSeek's V4.1 Flash news post of 10 September 2026 describes a 552B-parameter MoE with a Causal Encoder-Decoder architecture of 8B active parameters for input and 16B for output, says benchmark results are ahead of flagship models including DeepSeek-V4-Pro and that tests by multiple parties put it ahead of V4-Pro on performance, cost, speed and total runtime, and says from 04:00 UTC on 14 September 2026 deepseek-v4-pro requests route to V4.1-Flash at V4.1-Flash rates until V4.1-Pro launches. The Hugging Face model card snippet shows license mit, and an arXiv paper (2609.19969) covers its KV cache compression.

How it compares

552 billion is first-party. MIT rests on the model card snippet and the license file was not read. Only 8 billion active per token is incomplete: the post gives 8B for input and 16B for output, and an earlier note quoted the 16B half. Reported September 14 is not the release report, since the post is dated 10 September and 14 September is when V4-Pro traffic began routing to V4.1-Flash. Beating V4-Pro on agentic coding is DeepSeek's own vendor claim; the benchmark charts are images that were not read and the multiple parties were not identified. The race moved to sparse math is the author's framing.

Related work

Watch next

  • The benchmark tables, the identity of the third-party tests and V4.1-Pro.

Sources

  1. DeepSeek V4.1 Flash (DeepSeek news, 10 Sep 2026)deepseek.com
  2. DeepSeek-V4.1-Flash (Hugging Face)huggingface.co
  3. DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression (arXiv 2609.19969)arxiv.org

Provenance

The note above is reproduced unedited from the original post, first published on Threads on 21 September 2026 at 10:47 IST. Sources are the papers and datasets the note draws on.

View the original post
Embed this note
<iframe src="https://founder.myndlabs.tech/notes/embed/the-frontier-open-model-race-just-moved-to-DdiY-PoAFNb" width="480" height="420" style="border:0;max-width:100%" loading="lazy" title="The frontier open model race just moved to sparse math. deepseek v4.1 flash, reported september 14,…"></iframe>

More notes