DeepSeek V4 Flash
deepseek/deepseek-v4-flashVery low cost per token, with peak-hour pricing at 2x.
Model overview
Context window
1.048576M
tokens
Input price
R4.60
per 1M tokens
Output price
R13.70
per 1M tokens
Weekly volume
—
tokens / week
Open weights
No
API only
Capabilities
3
of 11 tags
Description
DeepSeek V4 Flash is a language model from DeepSeek. Webparam does not stock this model yet and can source it on request — tell us what you need it for and we will come back with availability and a price. DeepSeek’s stated focus is open-weight frontier reasoning at disruptive economics. The weights are not published, so it is API-only.
Capabilities
- Tool calling
- Calls functions you define and returns their arguments as structured data, the basis of agents.
- JSON mode
- Constrains the response to valid JSON matching a schema you supply.
- Streaming
- Emits tokens as they are generated, so answers appear progressively rather than all at once.
Strengths & weaknesses
Strengths
- Cheapest per token on the rate card
- Open weights permit self-hosting later
- Price quoted from the supplier's own rate card, not estimated
Weaknesses
- Priced at 2x during the supplier's peak hours
- English long-form prose trails Western flagships
- Not provisioned on this account yet — lead time applies
Pricing
| Rate | Price | Unit |
|---|---|---|
| Input | R4.60 | per 1M tokens |
| Output | R13.70 | per 1M tokens |
This model is not provisioned yet; the rate is quoted on request.
Context window
1.048576Mtokens
covers prompt and response together, so a long input leaves less room for the answer.
Supported features
| Feature | Support |
|---|---|
| Tool calling | Supported |
| JSON mode | Supported |
| Streaming | Supported |
| Vision input | Not supported |
| Audio input | Not supported |
| Long context | Not supported |
| Open weights | Not supported |
| Reasoning | Not supported |
| Fine-tunable | Not supported |
| Batch | Not supported |
| Caching | Not supported |
Example use cases
General assistants
Retrieval-augmented answering
Everyday production traffic
Comparison
Compare against anything →| Attribute | DeepSeek V4 Flash | DeepSeek V4 Pro | AC Flash 4.5 |
|---|---|---|---|
| Context window | 1.048576M | 1.048576M | — |
| Input price | R4.60 | R13.70 | R10.00 |
| Output price | R13.70 | R41.30 | R40.00 |
| Tool calling | Supported | Supported | Supported |
| JSON mode | Supported | Supported | Supported |
| Vision input | Not supported | Not supported | Not supported |
| Open weights | Not supported | Not supported | Not supported |
| Prompt caching | Not supported | Not supported | Not supported |
Best value in each row is highlighted. Blank cells are facts the supplier does not publish, not zeroes.
Provider information
View DeepSeek →DeepSeek · 深度求索 · Hangzhou, China · est. 2023
Spun out of a quantitative trading firm, DeepSeek built its reputation by releasing open-weight models that matched closed frontier quality at a fraction of the price. Its V-series redefined cost expectations for production LLMs, and its R-series brought open reasoning models into serious enterprise conversations.
Recent releases
Documentation
Frequently asked questions
Input is billed at R4.60 per 1M tokens and output at R13.70 per 1M tokens. Every figure on this page is an illustrative placeholder.