Skip to content
Dashboard

Deepseek V4 Flash

Deepseek V4 Flash is DeepSeek's April 23, 2026 efficiency-tier model in the V4 series. It pairs a hybrid attention architecture with a context window of 1.0M tokens and supports reasoning, tool use, and implicit caching. Your use is subject to DeepSeek's Terms & Privacy Policies.

ReasoningTool UseImplicit Caching
import { streamText } from 'ai'
const result = streamText({
model: 'deepseek/deepseek-v4-flash',
prompt: 'Why is the sky blue?'
})
Read docs

Providers

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.

Provider
Context
Max Output
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
ZDR
No Training
Release Date
262K52K
1.5s
162tps
$0.11/M
$0.22/M
Read:$0.03/M
Write:
04/23/2026
1M384K
1.6s
105tps
$0.14/M
$0.28/M
Read:$0.0/M
Write:
04/23/2026
Novita AI
90% off
Legal:TermsPrivacy
1M393K
1.5s
106tps
$0.14/M$0.01/M
$0.28/M$0.03/M
Read:$0.03/M$0.0/M
Write:
04/23/2026
1M1M
2.2s
33tps
$0.09/M
$0.18/M
Read:$0.02/M
Write:
04/23/2026
1M384K
2.7s
130tps
$0.14/M
$0.28/M
Read:$0.03/M
Write:
04/23/2026
1M128K
4.2s
99tps
$0.19/M
$0.51/M
Read:$0.03/M
Write:
04/23/2026
1M384K
3.7s
148tps
$0.20/M
$0.40/M
Read:$0.04/M
Write:
04/23/2026
1M384K
2.3s
312tps
$0.13/M
$0.26/M
Read:$0.03/M
Write:
04/23/2026
1M1M
0.4s
468tps
$0.28/M
$0.56/M
Read:$0.07/M
Write:
04/23/2026
1M384K
4.2s
151tps
$0.13/M
$0.25/M
Read:$0.03/M
Write:
04/23/2026