DeepSeek launches lightweight V4.1 Flash model
DeepSeek introduced V4.1 Flash, a 552 billion parameter mixture of experts model that activates only a small portion of its parameters per pass. The lightweight KV cache footprint enables it to handle complex agentic tasks efficiently.