Wednesday, July 29, 2026 · No. 367Prices synced 4 hours ago
All·AI·Model

The AI Model Price Index

Llama 4 Maverick 17B Instruct

by Charm Hyper · family llama · listed Apr 2026TW

Open multimodal Llama for strong reasoning with efficient everyday serving

Current rates USD per 1M tokens

Input$0.284
Output$0.934
Cache read
Cache write$0.142

Against the market

Input price vs. maker flagships ($/1M tokens)
DeepSeek V4 FlashDeepSeek V4 Flash: $0.14 per 1M tokens$0.14Llama 4 Maverick 17B InstructLlama 4 Maverick 17B Instruct: $0.284 per 1M tokens$0.284Gemini 3.5 Flash LiteGemini 3.5 Flash Lite: $0.30 per 1M tokens$0.30Qwen3.7 PlusQwen3.7 Plus: $0.50 per 1M tokens$0.50Mistral Medium (latest)Mistral Medium (latest): $1.50 per 1M tokens$1.50Grok 4.5Grok 4.5: $2.00 per 1M tokens$2.00Claude Opus 5Claude Opus 5: $5.00 per 1M tokens$5.00GPT-5.6 SolGPT-5.6 Sol: $5.00 per 1M tokens$5.00
Output price vs. maker flagships ($/1M tokens)
DeepSeek V4 FlashDeepSeek V4 Flash: $0.28 per 1M tokens$0.28Llama 4 Maverick 17B InstructLlama 4 Maverick 17B Instruct: $0.934 per 1M tokens$0.934Gemini 3.5 Flash LiteGemini 3.5 Flash Lite: $2.50 per 1M tokens$2.50Qwen3.7 PlusQwen3.7 Plus: $3.00 per 1M tokens$3.00Grok 4.5Grok 4.5: $6.00 per 1M tokens$6.00Mistral Medium (latest)Mistral Medium (latest): $7.50 per 1M tokens$7.50Claude Opus 5Claude Opus 5: $25.00 per 1M tokens$25.00GPT-5.6 SolGPT-5.6 Sol: $30.00 per 1M tokens$30.00

On the record

Model id
llama-4-maverick-17b-128e-instruct-fp8
Context window
430K tokens
Max output
43K tokens
Input modalities
text
Output modalities
text
Reasoning
no
Tool calls
yes
Open weights
yes
Knowledge cutoff
Aug 2024
Released
2026-04-30
Last updated
2026-07-22
Provider docs
hyper.charm.land

Back of the envelope

Also listed at

← All Charm Hyper listings