deepseek-v4-flash

mit

DeepSeek-V4-Flash hybrid model with both non-thinking and thinking (default) modes.

Context
1M
Max output
384K
Input price
$0.01/1M tokens
Output price
$0.03/1M tokens

Capabilities

vision
tool call
structured output
reasoning
json mode
streaming
fine tuning
batch

Details

Model IDdeepseek/deepseek-v4-flash
Creator DeepSeek
Familydeepseek
Licensemit
Parameters
Statusactive
Input modalitiestext
Output modalitiestext
Architecture
Knowledge cutoff
Training data cutoff
Release date
Deprecation date
Typechat
Reasoning tokensYes
Max input
Open weightYes
Sourceofficial
Last updated

Tools

Function Callingfunction_calling
Call external functions and APIs

Pricing

Input
$0.01
↓89%$0.09
Output
$0.03
↓83%$0.18
Cache write
Cache read
Batch in
Batch out
Price history · $/1M input
$0.14
$0.09
$0.01
Apr 25Jul 17Aug 5
DateInputOutputChange
Apr 25, 2026$0.14$0.28
Jul 17, 202620c73ac$0.09$0.18↓36%
Aug 5, 202619c9fe3$0.01$0.03↓89%

Input is ↓93% since Apr 25, 2026, from $0.14 to $0.01 per 1M tokens.

Price across providers

ProviderInputOutputSince launch
Vercel AI Gatewaycheapest$0.01$0.03↓93%
Novita AI$0.14$0.28
OpenRouter$0.14$0.28

Family Comparison: deepseek

ModelContextMax outPricing
deepseek-v4-flash-07311M$0.01$0.03
deepseek-v4-flash1M384K$0.01$0.03
deepseek-v3.1164K$0.25$0.95
deepseek-v3.2-thinking164K$0.26$0.38
deepseek-v3.1-terminus131K$0.27$1
deepseek-v3164K$0.27$1.12
deepseek-v3.2164K$0.28$0.42
deepseek-v4-pro1M384K$0.43$0.87

API

GET/v1/models/vercel/deepseek/deepseek-v4-flash