Model introduction
DeepSeek V4-Flash is the fast, economical open-weight V4 model. It supports a 1M context window, thinking and non-thinking modes, and OpenAI- and Anthropic-compatible APIs.
Model capabilities
Specifications come from the cited model material. Catalog tags help identify access features; use the source documentation for exact limits.
Verified specifications
- Official positioning
- Fast, efficient, economical V4 model
- Context window
- 1,000,000 tokens
- Maximum output
- 384,000 tokens
- Model size
- 284B total / 13B active parameters
Model pricing and access
Prices are read directly from the MoleAPI console catalog and shown by current billing group and context tier.
Your final charge follows the account group shown in the console.
Live pricing source
Pricing table
USD / 1M tokensStandardx1defaultInput $0.2 · Output $0.67
Input
$0.2 / 1M tokens
Output
$0.67 / 1M tokens
Cache read
$0.0075 / 1M tokens
Cache write
$0.25 / 1M tokens
Access protocols
Available billing groups: Standard
- openai
- POST
- /v1/chat/completions
About deepseek-v4-flash-260425
This introduction is transcreated for clarity and cross-checked against the cited model material.
DeepSeek V4-Flash is the fast, economical open-weight V4 model. It supports a 1M context window, thinking and non-thinking modes, and OpenAI- and Anthropic-compatible APIs.
Core strengths
- Independent testing shows much faster output than V4-Pro while retaining strong open-weight capability.
- A 1M context window and up to 384K output fit long inputs and long results.
- The 284B-total, 13B-active architecture is served at a low official API price.
Limitations
- Composite capability trails V4-Pro, making it a throughput-first choice rather than the highest-quality V4.
- Maximum reasoning effort can be very verbose, so real cost still depends on output token use.
Selection and production evaluation
Start a deepseek-v4-flash-260425 evaluation by mapping its official positioning to real work: High-throughput coding, agent sub-tasks, long-document processing, and cost-sensitive reasoning workloads. The first pass should exercise both its main strength, "Independent testing shows much faster output than V4-Pro while retaining strong open-weight capability.", and its known limitation, "Composite capability trails V4-Pro, making it a throughput-first choice rather than the highest-quality V4.", instead of relying on a single subjective general-chat comparison.
For access, MoleAPI currently lists openai protocols and the Standard billing groups for this DeepSeek model; the default price summary is Input $0.2 / 1M tokens · Output $0.67 / 1M tokens. Pricing, protocols, and groups come from the live catalog, so production planning should still price representative requests using real context length, output size, and cache-hit assumptions.
No independent ranking is shown unless it matches this exact model ID and reasoning profile, so nearby variants are not used as a proxy. Before launch, pin the model ID, prompt, and sample set, then compare task accuracy, structured-output validity, tool-call success, and timeout rates under the same conditions.
The model material on this page was last checked on 2026-07-23. When upstream model cards, context limits, or tool support change, update the cited bilingual facts before changing the recommendation; live MoleAPI price changes remain separate and update from the catalog automatically.
Code examples
These examples use MoleAPI's OpenAI-compatible endpoint and run after you replace the API key.
curl https://api.moleapi.com/v1/chat/completions \
-H "Authorization: Bearer $MOLEAPI_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek-v4-flash-260425",
"messages": [{"role": "user", "content": "Hello"}]
}'Frequently asked questions
How is deepseek-v4-flash-260425 priced?
This page reads prices from the MoleAPI console API and updates with model prices, context tiers, and account groups.
How can I access deepseek-v4-flash-260425?
Protocols and endpoints come from the supported_endpoint_types field in the MoleAPI model catalog.
How do I switch an existing project to deepseek-v4-flash-260425?
Keep the MoleAPI API address and key, replace the model parameter with the model ID on this page, then check protocol-specific parameter differences.
Where do the deepseek-v4-flash-260425 model details come from?
Capabilities and limitations are checked against DeepSeek and the other cited pages. Pricing and available protocols come only from the MoleAPI console.