DeepSeek Launches V4.1-Flash: 552B Params, 1M-Token Context — SkimNews

Get the Tech newsletter
Daily tech — startups, AI labs, chips, the launches that shape the next decade. Free.
- DeepSeek launched DeepSeek-V4.1-Flash on Thursday, describing it as the smallest model in its lineup built on the company's new Causal Encoder-Decoder architecture.
- DeepSeek-V4.1-Flash features 552B backbone parameters and supports a 1M-token context window.
Why it matters: DeepSeek's smallest model in the V4.1 family still carries 552B parameters and a 1M-token context window, giving the Chinese startup a lighter-weight option that still debuts its new Causal Encoder-Decoder architecture.
Ask SkimNews


