Skip to content
ModelPriceLab

Search model prices

Find source-linked rates by model, vendor, or platform.

Changes/DeepSeek launched V4 Pro and V4 Flash with thinking mode controls
MajorDeepSeekNew Launchdeepseek-v4-proVerified

DeepSeek launched V4 Pro and V4 Flash with thinking mode controls

DeepSeek added deepseek-v4-pro and deepseek-v4-flash to its API model catalog on April 24, 2026. Both models support thinking and non-thinking modes, 1M token context, 384K max output, JSON output, tool calls, and FIM in non-thinking mode. Official pricing lists V4 Flash at $0.14 cache-miss input / $0.028 cache-hit input / $0.28 output per 1M tokens, and V4 Pro at $1.74 cache-miss input / $0.145 cache-hit input / $3.48 output per 1M tokens.

Apr 24, 2026 (UTC)Views:20Effective: Apr 24, 2026 (UTC)

Impact Assessment

Cost5
Reliability2
Migration3
Quality5
Compliance1

Sources

social
DeepSeek V4 announcement on X

Official DeepSeek social announcement URL for the V4 Pro and V4 Flash launch.

pricing
DeepSeek Models & Pricing

DeepSeek lists deepseek-v4-flash and deepseek-v4-pro with 1M context, 384K max output, cache-hit input, cache-miss input, and output token pricing.

docs
DeepSeek Thinking Mode

DeepSeek documents thinking mode toggles, reasoning effort controls, reasoning_content behavior, and tool-call handling for deepseek-v4-pro.

model_card
DeepSeek V4 Pro model card

DeepSeek describes V4 Pro and V4 Flash as MoE models with 1M token context and configurable non-think, high, and max reasoning modes.

Comments

Loading comments...
AI Cost Optimization

Does this change your best option?

Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.