deepseek-v4-flash

Common Name: DeepSeek V4 Flash

DeepSeek
Released on Jul 31 12:00 AMSupportedTool InvocationSupportedReasoning
CompareTry in Chat

DeepSeek V4 Flash is a hybrid-thinking model with 1M context and 384K max output. It supports both non-thinking and thinking modes, with thinking enabled by default.

Specifications

Context
1000K
Maximum Output
384K
Inputtext
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Pricing

Peak window (local): 01:00–04:00, 06:00–10:00

Off-peak
Input¥1.50/MTokens
Cached Input¥0.05/MTokens
Output¥4.50/MTokens
Peak
Input¥3.00/MTokens
Cached Input¥0.10/MTokens
Output¥9.00/MTokens

Performance Metrics (24h)

Similar Models

¥1.50/¥4.50/M
ctx1.0Mmax384Kavailtps
InOutCap

DeepSeek V4 Flash Vision Exp is the experimental vision variant of V4 Flash. Text capability matches the stable V4 Flash, and images are billed as token-equivalent. Supports 1M context and 384K max output.

¥4.50/¥13.50/M
ctx1.0Mmax384Kavailtps
InOutCap

DeepSeek V4 Pro is the higher-capability hybrid-thinking model in the V4 family. It supports both non-thinking and thinking modes, with 1M context and 384K max output.

$2.20/$13.20/M
ctx1.1Mmax128Kavailtps
InOutCap

GPT-5.6 Terra balances GPT-5.6 intelligence and cost for strong performance at lower price points.

$2.75/$16.50/M
ctx1.1Mmax128Kavailtps
InOutCap

GPT-5.4 is OpenAI's most capable frontier model for professional work, unifying GPT and Codex lines. Features native computer use, 1M+ context, and configurable reasoning effort (none to xhigh).