DeepSeekLaunchAI ModelsSeptember 10, 2026

DeepSeek releases V4.1-Flash with new YOCO decoder-decoder architecture

DeepSeek-V4.1-Flash packs 552B backbone parameters, native image and text input, and up to a 1M-token context window. The YOCO architecture cuts GPU memory and prefill latency; Bloomberg reports pricing as low as a fraction of a cent per million tokens.

How this story unfolded

4 days · 10 reports · 36 community posts · 46 of 49 shown

  1. Sep 9
  2. Sep 10
  3. Sep 11
  4. Sep 12
  5. Sep 13

More stories today

Open the live feed