DeepSeek V4.1 Flash
by Requesty · family deepseek-flash · listed Sep 2026RTVW
DeepSeek V4.1 Flash model for reasoning and agentic coding
Current rates USD per 1M tokens
Input$0.22
Output$0.66
Cache read$0.007
Cache write—
Against the market
On the record
- Model id
deepseek-v4.1-flash- Context window
- 1.0M tokens
- Max output
- 393K tokens
- Input modalities
- text, image
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- May 2025
- Released
- 2026-09-10
- Last updated
- 2026-09-10
- Provider docs
- requesty.ai/solution/llm-routing/models
Back of the envelope
≈
Also listed at
- NanoGPT$0.15 in · $0.60 out
- LLM Gateway$0.15 in · $0.60 out
- DevPass (LLM Gateway)$0.15 in · $0.60 out
- Merge Gateway$0.15 in · $0.60 out
- OpenRouter$0.15 in · $0.60 out
- ClinePass$0.15 in · $0.60 out
- OpenCode Go$0.15 in · $0.60 out
- AIHubMix$0.155 in · $0.62 out
- Deep Infra$0.20 in · $0.60 out
- Eden AI$0.20 in · $0.60 out
- GreenPT$0.256 in · $1.28 out
- CrossModel$0.27 in · $1.08 out
- Charm Hyper$0.30 in · $1.20 out
- Hugging Face$0.30 in · $1.20 out
- Baseten$0.30 in · $1.20 out