DeepSeek V4 Flashdeepseek/deepseek-v4-flash
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
Pricing across providers
| Provider | Input /1M | Output /1M | Blended /1M | Latency p50 | Format | Freshness | Action |
|---|
1 hosting hidden — prices older than 60 days are pending re-verification.
Affiliate disclosure: We may earn a commission from qualified signups. Pricing independence is enforced at the data layer — see our Editorial Independence Policy.
Works with
Point any of these clients at a hosting's base URL — they all speak at least one of this model's endpoint protocols (OPENAI_COMPATIBLE).
Capabilities
- chat
- code
Languages: en, zh
Code samples
Example using DeepSeek — the cheapest hosting for this model as of last verification. Swap base_url and model to use a different provider from the matrix above.
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://api.deepseek.com/v1",
)
response = client.chat.completions.create(
model="deepseek-v4-flash",
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
Technical specs
- Context
- 1049K
- Max output
- 384K
- Parameters
- —
- Release
- —
- Training cutoff
- —
- License
- —
Similar models
Compare with
- DeepSeek V4 Flash vs DeepSeek V3.1Comparison planned — not yet published
- DeepSeek V4 Flash vs DeepSeek V4 ProComparison planned — not yet published
- DeepSeek V4 Flash vs GLM 5.1Comparison planned — not yet published
Frequently asked
How much does DeepSeek V4 Flash cost?+−
How do I access DeepSeek V4 Flash from outside China?+−
Is DeepSeek V4 Flash open-source? Can I fine-tune it?+−
Is DeepSeek V4 Flash OpenAI-compatible?+−
openai SDK client at the Provider's base_url and use the Provider's model name. See the Code Samples above for a copy-pasteable example.