DocumentationTry Hanzo

DeepSeek Models via Hanzo AI

15 models available

Access all 15 DeepSeek models through Hanzo's OpenAI-compatible API. One key, one bill, no rate limit juggling.

DeepSeek: DeepSeek V3
164K

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...

text
DeepSeek: DeepSeek V3 0324
164K

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the DeepSeek V3 model and performs really well...

text
DeepSeek: DeepSeek V3.1
164K

DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...

text
DeepSeek: R1
64K

DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

text
DeepSeek: R1 0528
164K

May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

text
DeepSeek: R1 Distill Llama 70B
8K

DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1. The model combines advanced distillation techniques to achieve high performance across...

text
DeepSeek: DeepSeek V3.1 Terminus
164K

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...

text
DeepSeek: DeepSeek V3.2
164K

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...

text
DeepSeek: DeepSeek V3.2 Exp
164K

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...

text
DeepSeek: DeepSeek V4 Flash 0423
1M

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

text
DeepSeek: DeepSeek V4 Flash 0731
1M

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

text
DeepSeek: DeepSeek V4 Flash Vision Exp
1M

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of DeepSeek V4 Flash 0731 from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

text
DeepSeek: DeepSeek V4 Pro 0423
1M

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

text
DeepSeek: DeepSeek V4 Pro 0813
1M

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

text
DeepSeek V4 Flash Latest
1M

This model always redirects to the latest model in the DeepSeek V4 Flash family.

text

Use DeepSeek models via Hanzo

One key. One bill. Works with every SDK you already have.