Mistral Small 4
by Requesty · family mistral-small · listed Mar 2026RTVW
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Current rates USD per 1M tokens
Input$0.148
Output$0.594
Cache read$0.148
Cache write—
Against the market
On the record
- Model id
mistral-small-2603- Context window
- 256K tokens
- Max output
- 256K tokens
- Input modalities
- text, image
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- Jun 2025
- Released
- 2026-03-16
- Last updated
- 2026-03-16
- Provider docs
- requesty.ai/solution/llm-routing/models
Back of the envelope
≈
Also listed at
- Cortecs$0.143 in · $0.568 out
- OpenRouter$0.15 in · $0.60 out
- Mistral$0.15 in · $0.60 out
- Eden AI$0.15 in · $0.60 out
- Kilo Gateway$0.15 in · $0.60 out
- Venice AI$0.188 in · $0.75 out