Wednesday, September 23, 2026 · No. 423Prices synced 5 hours ago
All·AI·Model

The AI Model Price Index

Llama 4 Maverick 17B 128E Instruct FP8

by watsonx.ai · family llama · listed Apr 2025TVW

Open multimodal Llama for strong reasoning with efficient everyday serving

Current rates USD per 1M tokens

Input$0.371
Output$1.48
Cache read
Cache write

Against the market

Input price vs. maker flagships ($/1M tokens)
DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.15 per 1M tokens$0.15Qwen3.8 FlashQwen3.8 Flash: $0.15 per 1M tokens$0.15Llama 4 Maverick 17B 128E Instruct FP8Llama 4 Maverick 17B 128E Instruct FP8: $0.371 per 1M tokens$0.371Gemini 3.8 FlashGemini 3.8 Flash: $0.75 per 1M tokens$0.75GLM-5.3GLM-5.3: $1.40 per 1M tokens$1.40GPT-6 SolGPT-6 Sol: $2.00 per 1M tokens$2.00Grok 4.7Grok 4.7: $2.00 per 1M tokens$2.00Claude Opus 5.5Claude Opus 5.5: $4.00 per 1M tokens$4.00
Output price vs. maker flagships ($/1M tokens)
Qwen3.8 FlashQwen3.8 Flash: $0.47 per 1M tokens$0.47DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.60 per 1M tokens$0.60Llama 4 Maverick 17B 128E Instruct FP8Llama 4 Maverick 17B 128E Instruct FP8: $1.48 per 1M tokens$1.48Gemini 3.8 FlashGemini 3.8 Flash: $3.75 per 1M tokens$3.75GLM-5.3GLM-5.3: $4.40 per 1M tokens$4.40Grok 4.7Grok 4.7: $6.00 per 1M tokens$6.00GPT-6 SolGPT-6 Sol: $10.00 per 1M tokens$10.00Claude Opus 5.5Claude Opus 5.5: $20.00 per 1M tokens$20.00

On the record

Model id
meta-llama/llama-4-maverick-17b-128e-instruct-fp8
Context window
131K tokens
Max output
8K tokens
Input modalities
text, image
Output modalities
text
Reasoning
no
Tool calls
yes
Open weights
yes
Knowledge cutoff
Aug 2024
Released
2025-04-05
Last updated
2025-04-05

Back of the envelope

Also listed at

← All watsonx.ai listings