DeepSeek-V4.1-Flash debuts
DeepSeek launched DeepSeek-V4.1-Flash last night with a 552-billion-parameter mixture-of-experts backbone, native vision, a 1-million-token context window and an architecture built to make repeatedly reading large contexts cheaper.
Its open weights are are available for developers and enterprises to download and use for commercial purposes under a permissive, enterprise-friendly MIT License on Hugging Face. For developers evaluating the model for coding agents and other long-running workflows, however, the headline API rate is immediately enticing.