DeepSeek has released DeepSeek-V4.1-Flash, a 552-billion-parameter multimodal Mixture-of-Experts model built around a Causal Encoder–Decoder layout and a one-million-token context window. The company positions the model as the smallest member of a new architecture family aimed at higher capability ceilings, faster decoding, and larger deployment throughput, with API access under the model name deepseek-flash. The Causal Encoder–Decoder stack use…
This story is only covered by news sources that have yet to be evaluated by the independent media monitoring agencies we use to assess the quality and reliability of news outlets on our platform. Learn more here.