DeepSeek V4 Flash 0731
by Venice AI · family deepseek-flash · listed Jul 2026RTW
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Current rates USD per 1M tokens
Input$0.175
Output$0.35
Cache read$0.035
Cache write—
Against the market
On the record
- Model id
deepseek-v4-flash-0731- Context window
- 1M tokens
- Max output
- 33K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- May 2025
- Released
- 2026-07-31
- Last updated
- 2026-08-01
- Provider docs
- docs.venice.ai
Back of the envelope
≈
Also listed at
- Alibaba Token Planfree in · free out
- SCNet Token Planfree in · free out
- Alibaba Token Plan (China)free in · free out
- Deep Infra$0.08 in · $0.18 out
- Ambient$0.08 in · $0.18 out
- DigitalOcean$0.08 in · $0.252 out
- Eden AI$0.08 in · $0.18 out
- CrofAI$0.12 in · $0.21 out
- Requesty$0.126 in · $0.252 out
- Cortecs$0.13 in · $0.28 out
- Inceptron$0.13 in · $0.28 out
- Perplexity Agent$0.13 in · $0.26 out
- Baseten$0.13 in · $0.26 out
- Vercel AI Gateway$0.13 in · $0.26 out
- RunInfra$0.13 in · $0.27 out