Skip to content
ModelPriceLab

Search model prices

Find source-linked rates by model, vendor, or platform.

Model Services/DeepSeek/DeepSeek V4.1 Flash
deepseek

DeepSeek V4.1 Flash

Active

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

View current API pricing for this model
Input, output, and cached rates with source and verification time, plus same-workload cost examples.
Go to pricing page →
Input / 1M tokens
$0.0030
Output / 1M tokens
$2.40
Record updated 2026-10-05
Modality:LLM
Context:1,048,576 tokens

Pricing

Platform
Scope
Input
Output
Details
OpenRouter
Unified API Router
Standard
$0.0030
$2.40
1.0M context · 944k max output

How much are you overpaying?

Enter your monthly spend on this model to see what the same workload could cost on cheaper alternatives.

$
Enter a monthly spend to see the listed-rate comparison.

Latest changes

No recent changes.

Benchmarks

No benchmark data available.

⚡DeepSeek V4.1 Flash Comparisons & Scenario Guides

❓DeepSeek V4.1 Flash Frequently Asked Questions

How much does DeepSeek V4.1 Flash cost per 1M tokens?▼

DeepSeek V4.1 Flash is priced at $0.0030 per 1M input tokens and $2.40 per 1M output tokens, with cached input at $0.0030/1M on openrouter. Check provider terms before deployment.

What is the context window for DeepSeek V4.1 Flash?▼

DeepSeek V4.1 Flash supports up to 1,048,576 tokens context with a max completion limit of 943,718 tokens.

What are cost-saving alternatives to DeepSeek V4.1 Flash?▼

ModelPriceLab automatically calculates cheaper alternatives in the same model class, showing monthly spend differences and migration guides.

AI Cost Optimization

Does this change your best option?

Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.