Skip to content

Deepseek V4 1 Flash

Topic archive1 matches

Back to homeGEO summary endpoint

2026-09-20

Technology

  • DeepSeek releases DeepSeek-V4.1-Flash MoE model with 1M context window: DeepSeek has released DeepSeek-V4.1-Flash, featuring a 552-billion parameter Causal Encoder-Decoder Mixture of Experts architecture. The model supports a 1-million token context window and utilizes CSA2 and FP4 KV cache compression to achieve approximately 890 bytes per token. It is available via API as deepseek-flash.

    AI ModelsPandaily

    Permalink