DeepSeek V4 Flash 0731
by Weights & Biases · family deepseek · listed Jul 2026RTW
DeepSeek V4-Flash-0731 is an MoE model great for coding, reasoning, and agentic workloads.
Current rates USD per 1M tokens
Input$0.13
Output$0.28
Cache read$0.07
Cache write—
Against the market
On the record
- Model id
deepseek-ai/DeepSeek-V4-Flash-0731- Context window
- 262K tokens
- Max output
- 262K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- May 2025
- Released
- 2026-07-31
- Last updated
- 2026-07-31
- Provider docs
- docs.wandb.ai/guides/integrations/inference/
Back of the envelope
≈
Also listed at
- Alibaba Token Planfree in · free out
- Alibaba Token Plan (China)free in · free out
- DigitalOcean$0.08 in · $0.252 out
- OpenRouter$0.09 in · $0.18 out
- Deep Infra$0.09 in · $0.18 out
- CrofAI$0.12 in · $0.21 out
- Baseten$0.13 in · $0.26 out
- Fireworks AI$0.14 in · $0.28 out
- NanoGPT$0.14 in · $0.28 out
- Ambient$0.14 in · $0.28 out
- EmpirioLabs AI$0.14 in · $0.28 out
- Together AI$0.14 in · $0.28 out
- Requesty$0.14 in · $0.28 out
- Kilo Gateway$0.14 in · $0.28 out
- Hugging Face$0.14 in · $0.28 out