Llama 4 Maverick 17B Instruct
by Charm Hyper · family llama · listed Apr 2026TW
Open multimodal Llama for strong reasoning with efficient everyday serving
Current rates USD per 1M tokens
Input$0.284
Output$0.934
Cache read—
Cache write$0.142
Against the market
On the record
- Model id
llama-4-maverick-17b-128e-instruct-fp8- Context window
- 430K tokens
- Max output
- 43K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- no
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- Aug 2024
- Released
- 2026-04-30
- Last updated
- 2026-07-22
- Provider docs
- hyper.charm.land
Back of the envelope
≈
Also listed at
- Llamafree in · free out
- GitHub Modelsfree in · free out
- Abacus$0.14 in · $0.59 out
- IO.NET$0.15 in · $0.60 out
- Deep Infra$0.20 in · $0.80 out
- Azure Cognitive Services$0.25 in · $1.00 out
- Azure$0.25 in · $1.00 out
- NovitaAI$0.27 in · $0.85 out