DeepSeek models
DeepSeek publishes open-weight models that compete with far more expensive closed models, particularly on reasoning and code. Zeplik carries the V4 generation (Pro and Flash) and the V3.x line behind it.
All 16 DeepSeek models
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...
DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.
DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...
May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...
DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....
DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...
Pricing at a glance
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| DeepSeek V4 Flash Vision Exp | 1.0M | 2.4 | 7.3 | August 2026 |
| DeepSeek V4 Pro 0813 | 1.0M | 6.4 | 19.1 | August 2026 |
| DeepSeek V4 Pro 0813 (batch) | 1.0M | 14.5 | 43.6 | August 2026 |
| DeepSeek V4 Flash 0731 | 1.0M | 0.72 | 2.0 | July 2026 |
| DeepSeek V4 Flash 0731 (batch) | 1.0M | 1.5 | 3.1 | July 2026 |
| DeepSeek V4 Pro 0423 | 1.0M | 9.1 | 18.2 | April 2026 |
| DeepSeek V4 Flash 0423 | 1.0M | 0.93 | 1.9 | April 2026 |
| DeepSeek V3.2 | 164K | 3.0 | 4.4 | December 2025 |
| DeepSeek V3.2 Exp | 164K | 3.0 | 4.5 | September 2025 |
| DeepSeek V3.1 Terminus | 131K | 3.0 | 11.0 | September 2025 |
| DeepSeek V3.1 | 161K | 6.0 | 18.2 | August 2025 |
| R1 0528 | 164K | 5.5 | 23.7 | May 2025 |
| DeepSeek V3 0324 | 164K | 2.8 | 11.0 | March 2025 |
| R1 Distill Llama 70B | 8K | 8.8 | 8.8 | January 2025 |
| R1 | 64K | 7.7 | 27.5 | January 2025 |
| DeepSeek V3 | 164K | 3.5 | 9.8 | December 2024 |
Other providers
Explore Zeplik
- 465 AI skillsReady-to-run expert methods that run on any DeepSeek model, or any other model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.