Deepseek V4 Flash

mit

DeepSeek-V4-Flash is a lightweight model meticulously designed by DeepSeek to deliver the ultimate combination of lightning-fast response times and unmatched cost-effectiveness.

Context
1.0M
Max output
393K
Input price
$0.14/1M tokens
Output price
$0.28/1M tokens

Capabilities

vision
tool call
structured output
reasoning
json mode
streaming
fine tuning
batch

Details

Model IDdeepseek/deepseek-v4-flash
Provider Novita AI
Creator DeepSeek
Familydeepseek
Licensemit
Parameters
Statusactive
Input modalitiestext
Output modalitiestext
Architecture
Knowledge cutoff
Training data cutoff
Release date
Deprecation date
Typechat
Reasoning tokens
Max input
Open weightYes
Sourceofficial
Last updated

Tools

Function Callingfunction_calling
Call external functions and APIs

Endpoints

Chat CompletionsPOST
Generate chat responses with messageshttps://api.novita.ai/v3/openai/v1/chat/completions

Pricing

Input
$0.14
Output
$0.28
Cache write
Cache read
Batch in
Batch out

Price across providers

ProviderInputOutputSince launch
Vercel AI Gatewaycheapest$0.01$0.03↓93%
Novita AI$0.14$0.28
OpenRouter$0.14$0.28

Family Comparison: deepseek

ModelContextMax outPricing
DeepSeek-OCR 28K8K$0.03$0.03
DeepSeek-OCR8K8K$0.03$0.03
Deepseek V4 Flash 07311.0M393K$0.14$0.28
Deepseek V4 Flash1.0M393K$0.14$0.28
Deepseek V3.2164K66K$0.27$0.4
DeepSeek V3 0324164K66K$0.27$1.12
Deepseek V3.1 Terminus131K33K$0.27$1
DeepSeek V3.1131K33K$0.27$1
Deepseek V3.2 Exp164K66K$0.27$0.41
DeepSeek V3 (Turbo) 64K16K$0.4$1.3
Deepseek Prover V2 671B160K160K$0.7$2.5
DeepSeek V364K8K$0.89$0.89
DeepSeek V364K16K$0.89$0.89
Deepseek V4 Pro1.0M393K$1.6$3.2

API

GET/v1/models/novita/deepseek/deepseek-v4-flash

Deepseek V4 Flash

deepseek/deepseek-v4-flash

DeepSeek-V4-Flash is a lightweight model meticulously designed by DeepSeek to deliver the ultimate combination of lightning-fast response times and unmatched cost-effectiveness.

Changes · 1 entries
deepseek/deepseek-v4-flashcreate08f0ffeMay 11, 2026, 08:27 AM