deepseek-flash

Common Name: DeepSeek V4.1 Flash

DeepSeek
Released on 12:00 AMSupportedTool InvocationSupportedReasoning
CompareTry in Chat

DeepSeek V4.1 Flash: new architecture with native multimodal vision understanding, outperforming V4 Pro across performance, cost, speed and total time. 1M context and 384K max output.

Specifications

Context
1000K
Maximum Output
384K
Inputtext, image
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Availability Trend (24h)

Pricing

Peak window (local): 01:00–04:00, 06:00–10:00

Off-peak
Input¥1.00/MTokens
Cached Input¥0.02/MTokens
Output¥4.00/MTokens
Peak
Input¥2.00/MTokens
Cached Input¥0.04/MTokens
Output¥8.00/MTokens

Performance Metrics (24h)

Similar Models

¥1.00/¥4.00/M
ctx1.0Mmax384Kavailtps
InOutCap

DeepSeek V4 Flash is a hybrid-thinking model with 1M context and 384K max output. It supports both non-thinking and thinking modes, with thinking enabled by default.

¥1.00/¥4.00/M
ctx1.0Mmax384Kavailtps
InOutCap

DeepSeek V4 Flash Vision Exp is the experimental vision variant of V4 Flash. Text capability matches the stable V4 Flash, and images are billed as token-equivalent. Supports 1M context and 384K max output.

¥4.50/¥13.50/M
ctx1.0Mmax384Kavailtps
InOutCap

DeepSeek V4 Pro is the higher-capability hybrid-thinking model in the V4 family. It supports both non-thinking and thinking modes, with 1M context and 384K max output.

$1.80/$10.80/M-10%
ctx1.1Mmax128Kavailtps

GPT-5.6 Terra balances GPT-5.6 intelligence and cost for strong performance at lower price points.