Saturday, July 25, 2026 · No. 363Prices synced 5 hours ago
All·AI·Model

The AI Model Price Index

DeepSeek V4 Flash

by Weights & Biases · family deepseek · listed Apr 2026RTW

DeepSeek V4-Flash is an MoE model with 1M context length great for coding, reasoning, and agentic workloads.

Current rates USD per 1M tokens

Input$0.14
Output$0.28
Cache read$0.07
Cache write

Against the market

Input price vs. maker flagships ($/1M tokens)
DeepSeek V4 FlashDeepSeek V4 Flash: $0.14 per 1M tokens$0.14DeepSeek V4 FlashDeepSeek V4 Flash: $0.14 per 1M tokens$0.14Gemini 3.5 Flash LiteGemini 3.5 Flash Lite: $0.30 per 1M tokens$0.30Qwen3.7 PlusQwen3.7 Plus: $0.50 per 1M tokens$0.50Mistral Medium (latest)Mistral Medium (latest): $1.50 per 1M tokens$1.50Grok 4.5Grok 4.5: $2.00 per 1M tokens$2.00Claude Opus 5Claude Opus 5: $5.00 per 1M tokens$5.00GPT-5.6 SolGPT-5.6 Sol: $5.00 per 1M tokens$5.00
Output price vs. maker flagships ($/1M tokens)
DeepSeek V4 FlashDeepSeek V4 Flash: $0.28 per 1M tokens$0.28DeepSeek V4 FlashDeepSeek V4 Flash: $0.28 per 1M tokens$0.28Gemini 3.5 Flash LiteGemini 3.5 Flash Lite: $2.50 per 1M tokens$2.50Qwen3.7 PlusQwen3.7 Plus: $3.00 per 1M tokens$3.00Grok 4.5Grok 4.5: $6.00 per 1M tokens$6.00Mistral Medium (latest)Mistral Medium (latest): $7.50 per 1M tokens$7.50Claude Opus 5Claude Opus 5: $25.00 per 1M tokens$25.00GPT-5.6 SolGPT-5.6 Sol: $30.00 per 1M tokens$30.00

On the record

Model id
deepseek-ai/DeepSeek-V4-Flash
Context window
1.0M tokens
Max output
1.0M tokens
Input modalities
text
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
yes
Knowledge cutoff
May 2025
Released
2026-04-24
Last updated
2026-04-24

Back of the envelope

Also listed at

← All Weights & Biases listings