11 text
Open-source 1.6T-parameter MoE with 49B active, built for advanced reasoning and long-horizon agents.
Efficiency-optimized DeepSeek V4 at 284B total / 13B active for fast, high-throughput inference.
Open-source model with GPT-5 class reasoning performance.
Experimental DeepSeek V3.2 variant with DeepSeek Sparse Attention.
Large hybrid reasoning model supporting thinking and non-thinking modes.
Updated DeepSeek R1 with performance on par with OpenAI o1.
March 2024 iteration of the DeepSeek V3 685B MoE model.
Distilled reasoning model based on Llama 3.3 70B using DeepSeek R1 outputs.
Open-source reasoning model matching OpenAI o1 performance.
Original DeepSeek V3 — the model that put DeepSeek on the map.
Distilled reasoning model based on Qwen 2.5 32B using DeepSeek R1 outputs.