DeepSeek

DeepSeek
Models from DeepSeek

DeepSeek V3.2
Open-source Mixture-of-Experts LLM tuned for high-efficiency reasoning, coding, and general language tasks across long-form prompts.

DeepSeek V4 Flash
Lightweight MoE model with 284B total / 13B active parameters and native 1M context, tuned for low-latency, cost-effective high-concurrency use.

DeepSeek V4 Flash 0731
Post-trained 0731 release with major gains across coding, repository work, tool use, and full-stack tasks, plus a 1M context window.

DeepSeek V4 Pro
Flagship MoE LLM with 1.6T total / 49B active parameters and native 1M context for advanced math, logical inference, and specialized coding.

DeepSeek V4 Pro 0813
Official 0813 Pro release with major gains across coding, repository work, tool use, and agent tasks, plus hybrid thinking and a 1M context window.

DeepSeek V4.1 Flash
Lightweight flagship of DeepSeek’s new architecture with native image understanding, hybrid thinking, 1M context, and up to 384K output tokens.

Janus-Pro DeepSeek
Autoregressive framework on the Janus Pro 7B model that unifies multimodal understanding and image generation in one architecture.
