DeepSeek V4 Flash
by Weights & Biases · family deepseek · listed Apr 2026RTW
DeepSeek V4-Flash is an MoE model with 1M context length great for coding, reasoning, and agentic workloads.
Current rates USD per 1M tokens
Input$0.14
Output$0.28
Cache read$0.07
Cache write—
Against the market
On the record
- Model id
deepseek-ai/DeepSeek-V4-Flash- Context window
- 1.0M tokens
- Max output
- 1.0M tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- May 2025
- Released
- 2026-04-24
- Last updated
- 2026-04-24
- Provider docs
- docs.wandb.ai/guides/integrations/inference/
Back of the envelope
≈
Also listed at
- Kenarifree in · free out
- Alibaba Token Planfree in · free out
- Alibaba Token Plan (China)free in · free out
- UnoRouter$0.0625 in · $0.125 out
- Deep Infra$0.09 in · $0.18 out
- OpenRouter$0.0938 in · $0.188 out
- Pioneer$0.10 in · $0.20 out
- routing.run$0.112 in · $0.224 out
- GMI Cloud$0.112 in · $0.224 out
- CrofAI$0.12 in · $0.21 out
- Cortecs$0.133 in · $0.266 out
- Venice AI$0.138 in · $0.275 out
- Fireworks AI$0.14 in · $0.28 out
- OpenCode Go$0.14 in · $0.28 out
- Nvidia$0.14 in · $0.28 out