DeepSeek V4.1 Flash is now available on Huggingface!
552B parameters MoE model with a new Encoder-Decoder structure.
V4.1 Flash adopted a new pre-training method and underwent larger-scale reinforcement learning post-training.
The smallest model in DeepSeek new architecture family, with native visual understanding.
🚀 Introducing DeepSeek-V4.1-Flash: smarter, faster, more efficient. 🔹 Introducing the smallest model in our new architecture family, with native visual unders...