DeepSeekLaunchAI ModelsSeptember 10, 2026

DeepSeek releases V4.1-Flash with novel YOCO decoder-decoder architecture

DeepSeek-V4.1-Flash packs 552B backbone parameters, native image and text input, and up to a 1M-token context window. The YOCO architecture cuts GPU memory and prefill latency; Bloomberg reports pricing as low as a fraction of a cent per million tokens.

How this story unfolded

4 days · 8 reports · 35 community posts · 43 of 46 shown

  1. Sep 9
  2. Sep 10
  3. Sep 11
  4. Sep 12
  5. Sep 13

More stories today

Open the live feed