Models/DeepSeek

DeepSeek

DeepSeek models, pricing, and protocols

20 currently available models with pricing synced from the MoleAPI console.

20 models

Compare
D
deepseek-v4-flash-90%Official profile verified

DeepSeek V4-Flash is the fast, economical open-weight V4 model. It supports a 1M context window, thinking and non-thinking modes, and OpenAI- and Anthropic-compatible APIs.

View model
defaultx1
Input$0.22 / 1M tokens
Output$0.67 / 1M tokens
Cache read$0.0075 / 1M tokens
Cache write$0.275 / 1M tokens
discountx0.8
Input$0.176 / 1M tokens
Output$0.536 / 1M tokens
Cache read$0.006 / 1M tokens
Cache write$0.22 / 1M tokens
relayx0.3
Input$0.066 / 1M tokens
Output$0.201 / 1M tokens
Cache read$0.0023 / 1M tokens
Cache write$0.0825 / 1M tokens
tempx0.1
Input$0.022 / 1M tokens
Output$0.067 / 1M tokens
Cache read$0.0008 / 1M tokens
Cache write$0.0275 / 1M tokens
AA index 40openaiopenai-responseanthropicgeminiopenai-response-compactopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
D
deepseek-v4-pro-90%Official profile verified

DeepSeek V4-Pro is the Pro model in DeepSeek's V4 family. DeepSeek's update notes confirm access through both OpenAI Chat Completions and Anthropic interfaces.

View model
defaultx1
Input$0.67 / 1M tokens
Output$1.34 / 1M tokens
Cache read$0.02 / 1M tokens
Cache write$0.83 / 1M tokens
tempx0.1
Input$0.067 / 1M tokens
Output$0.134 / 1M tokens
Cache read$0.002 / 1M tokens
Cache write$0.083 / 1M tokens
AA index 43openaiopenai-responseopenai-response-compactanthropicgeminiopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
D

deepseek-v4-flash-vision-exp is a DeepSeek model available through MoleAPI with openai, openai-response, anthropic, gemini access. It fits OpenAI compatible, Responses API, Anthropic compatible, Gemini compatible workflows, with pricing synced from the MoleAPI console.

View model
defaultx1
Input$0.22 / 1M tokens
Output$0.67 / 1M tokens
Cache read$0.0075 / 1M tokens
Cache write$0.275 / 1M tokens
discountx0.8
Input$0.176 / 1M tokens
Output$0.536 / 1M tokens
Cache read$0.006 / 1M tokens
Cache write$0.22 / 1M tokens
relayx0.3
Input$0.066 / 1M tokens
Output$0.201 / 1M tokens
Cache read$0.0023 / 1M tokens
Cache write$0.0825 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleVisionPrompt cacheCompare
D
deepseek-v4-flash-260425Official profile verified

DeepSeek V4-Flash is the fast, economical open-weight V4 model. It supports a 1M context window, thinking and non-thinking modes, and OpenAI- and Anthropic-compatible APIs.

defaultx1
Input$0.22 / 1M tokens
Output$0.67 / 1M tokens
Cache read$0.0075 / 1M tokens
Cache write$0.275 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
D
deepseek-v4-pro-260425Official profile verified

DeepSeek V4-Pro is the Pro model in DeepSeek's V4 family. DeepSeek's update notes confirm access through both OpenAI Chat Completions and Anthropic interfaces.

View model
defaultx1
Input$0.67 / 1M tokens
Output$1.34 / 1M tokens
Cache read$0.02 / 1M tokens
Cache write$0.83 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
D
deepseek-v3.2-70%Official profile verified

DeepSeek V3.2 is a reasoning-first open model for agent scenarios, using sparse attention for efficient long context and supporting tools in thinking and non-thinking modes.

defaultx1
Input$0.29 / 1M tokens
Output$0.435 / 1M tokens
Cache read$0.145 / 1M tokens
discountx0.8
Input$0.232 / 1M tokens
Output$0.348 / 1M tokens
Cache read$0.116 / 1M tokens
relayx0.3
Input$0.087 / 1M tokens
Output$0.1305 / 1M tokens
Cache read$0.0435 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
D

DeepSeek model by DeepSeek for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.3 / 1M tokens
Output$0.45 / 1M tokens
openaiOpenAI compatibleCompare
D

DeepSeek model by DeepSeek for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.56 / 1M tokens
Output$1.68 / 1M tokens
Cache read$0.28 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
D

DeepSeek model by DeepSeek for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.58 / 1M tokens
Output$1.74 / 1M tokens
Cache read$0.464 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
D

DeepSeek model by DeepSeek for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.29 / 1M tokens
Output$1.16 / 1M tokens
Cache read$0.0362 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
D

DeepSeek model by DeepSeek for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.56 / 1M tokens
Output$1.68 / 1M tokens
openaiOpenAI compatibleCompare
D
deepseek-v3-2-251201Official profile verified

DeepSeek V3.2 is a reasoning-first open model for agent scenarios, using sparse attention for efficient long context and supporting tools in thinking and non-thinking modes.

defaultx1
Input$0.29 / 1M tokens
Output$0.435 / 1M tokens
Cache read$0.145 / 1M tokens
openaiOpenAI compatibleCompare
D

DeepSeek model by DeepSeek for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.192 / 1M tokens
Output$0.192 / 1M tokens
openaiOpenAI compatibleCompare
D

DeepSeek reasoning model by DeepSeek for complex problem solving, analysis, coding, and agent workflows.

defaultx1
Input$0.58 / 1M tokens
Output$2.32 / 1M tokens
openaiOpenAI compatibleReasoningCompare
D

DeepSeek reasoning model by DeepSeek for complex problem solving, analysis, coding, and agent workflows.

defaultx1
Input$0.574 / 1M tokens
Output$2.294 / 1M tokens
Cache read$0.1148 / 1M tokens
openaiOpenAI compatibleReasoningPrompt cacheCompare
D

DeepSeek reasoning model by DeepSeek for complex problem solving, analysis, coding, and agent workflows.

defaultx1
Input$1 / 1M tokens
Output$3 / 1M tokens
openaiOpenAI compatibleReasoningCompare
D

DeepSeek reasoning model by DeepSeek for complex problem solving, analysis, coding, and agent workflows.

defaultx1
Input$0.6 / 1M tokens
Output$2.4 / 1M tokens
openaiOpenAI compatibleReasoningCompare
D

DeepSeek reasoning model by DeepSeek for complex problem solving, analysis, coding, and agent workflows.

defaultx1
Input$0.7 / 1M tokens
Output$2.5669 / 1M tokens
Cache read$0.35 / 1M tokens
openaiOpenAI compatibleReasoningPrompt cacheCompare
D

DeepSeek reasoning model by DeepSeek for complex problem solving, analysis, coding, and agent workflows.

defaultx1
Input$75 / 1M tokens
Output$75 / 1M tokens
openaiOpenAI compatibleReasoningCompare
D

DeepSeek vision model by DeepSeek for document parsing, OCR, tables, charts, and layout understanding.

defaultx1
Input$0 / 1M tokens
Output$0 / 1M tokens
openaiOpenAI compatibleCompare