ModelsDeepSeek V4 Flash (0731)
DeepSeek V4 Flash (0731)
deepseek-ai/DeepSeek-V4-Flash-0731DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.
ThinkingToolsRAG
✨New
Context
128K
Speed
Fast
Capabilities
Chat Completions
VisionNo
ThinkingYes
ToolsYes
InternetNo
RAGYes
InputText
OutputText
Parameters
| Parameter | Values | Recommended | |
|---|---|---|---|
| Temperature | min0 | max1 | 0.8 |
| Top P | min0 | max1 | 0.9 |
| Top K | min1 | max100 | 40 |
| Freq. Penalty | min-2 | max2 | 0 |
| Pres. Penalty | min-2 | max2 | 0 |
| Rep. Penalty | min1 | max2 | 1 |
| Max Tokens | min1 | max1024 | 512 |
Credit Cost per Plan (16K tokens context)
Cost scales with the context length you use.
Context Length
Showing cost for requests up to 16,384 tokens.
FreeN/A
Basic1 cr/req
≈ 1,000 req/month
Starter1 cr/req
≈ 2,000 req/month
Pro1 cr/req
≈ 3,000 req/month
Pro+1 cr/req
≈ 5,000 req/month
Max1 cr/req
≈ 12,000 req/month
Max+1 cr/req
≈ 25,000 req/month
Ultimate1 cr/req
≈ 50,000 req/month
Context Window
128K
131,072 tokens