Mistral Small 4
by Requesty · family mistral-small · listed Mar 2026RTVW
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Current rates USD per 1M tokens
Input$0.165
Output$0.66
Cache read$0.165
Cache write—
Against the market
On the record
- Model id
mistral-small-2603- Context window
- 256K tokens
- Max output
- 256K tokens
- Input modalities
- text, image
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- Jun 2025
- Released
- 2026-03-16
- Last updated
- 2026-03-16
- Provider docs
- requesty.ai/solution/llm-routing/models
Back of the envelope
≈
Also listed at
- LLM Gateway$0.15 in · $0.60 out
- Kilo Gateway$0.15 in · $0.60 out
- OpenRouter$0.15 in · $0.60 out
- Eden AI$0.15 in · $0.60 out
- Tempr Gateway$0.15 in · $0.60 out
- Mistral$0.15 in · $0.60 out
- DevPass (LLM Gateway)$0.15 in · $0.60 out
- Cortecs$0.156 in · $0.625 out
- Opper$0.581 in · $2.44 out