ModelsDeepSeek V4 Flash (0731)

DeepSeek V4 Flash (0731)

deepseek-ai/DeepSeek-V4-Flash-0731

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

ThinkingToolsRAG
New
Context

128K

Speed

Fast

Capabilities

Chat Completions
VisionNo
ThinkingYes
ToolsYes
InternetNo
RAGYes
InputText
OutputText

Parameters

ParameterValuesRecommended
Temperaturemin0max10.8
Top Pmin0max10.9
Top Kmin1max10040
Freq. Penaltymin-2max20
Pres. Penaltymin-2max20
Rep. Penaltymin1max21
Max Tokensmin1max1024512

Credit Cost per Plan (16K tokens context)

Cost scales with the context length you use.

Context Length

Showing cost for requests up to 16,384 tokens.

FreeN/A
Basic1 cr/req

1,000 req/month

Starter1 cr/req

2,000 req/month

Pro1 cr/req

3,000 req/month

Pro+1 cr/req

5,000 req/month

Max1 cr/req

12,000 req/month

Max+1 cr/req

25,000 req/month

Ultimate1 cr/req

50,000 req/month

Context Window

128K

131,072 tokens

Get Started

Subscribe to a plan to use this model via the API.

All Models