Introducing DeepSeek-V4.1-Flash: smarter, faster, more efficient.
Introducing the smallest model in our new architecture family, with native visual understanding. Designed for greater capability, faster inference, and higher throughput.
We introduce DeepSeek-V4.1-Flash, a multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens. The model natively processes images and text, and generates text autoregressively.
DeepSeek
DeepSeek-V4-Pro is a 1.6T-parameter Mixture-of-Experts model with 49B activated, supporting a one-million-token context. The 0813 revision is its GA release.
DeepSeek
DeepSeek-V4-Flash-0731 is the official release of DeepSeek-V4-Flash, superseding the preview version, with substantially enhanced agentic capabilities.
An occasional email when notable AI dev tools and models land in the directory. No spam, unsubscribe anytime.