Monday, October 5, 2026 · No. 435Prices synced 20 hours ago
All·AI·Model

The AI Model Price Index

Llama 4 Maverick 17b 128e Instruct

by Nvidia · listed Apr 2025TVW

Open multimodal Llama model for strong reasoning and fast responses

Current rates USD per 1M tokens

Inputfree
Outputfree
Cache read—
Cache write—

Against the market

Input price vs. maker flagships ($/1M tokens)
Llama 4 Maverick 17b 128e InstructLlama 4 Maverick 17b 128e Instruct: free per 1M tokensfreeDeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.15 per 1M tokens$0.15Qwen3.8 Omni FlashQwen3.8 Omni Flash: $0.15 per 1M tokens$0.15Gemini 3.8 FlashGemini 3.8 Flash: $0.75 per 1M tokens$0.75GLM-5.3GLM-5.3: $1.40 per 1M tokens$1.40Claude Sonnet 5.5Claude Sonnet 5.5: $2.00 per 1M tokens$2.00GPT-6.1 SolGPT-6.1 Sol: $2.00 per 1M tokens$2.00Grok 4.7Grok 4.7: $2.00 per 1M tokens$2.00
Output price vs. maker flagships ($/1M tokens)
Llama 4 Maverick 17b 128e InstructLlama 4 Maverick 17b 128e Instruct: free per 1M tokensfreeQwen3.8 Omni FlashQwen3.8 Omni Flash: $0.47 per 1M tokens$0.47DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.60 per 1M tokens$0.60Gemini 3.8 FlashGemini 3.8 Flash: $3.75 per 1M tokens$3.75GLM-5.3GLM-5.3: $4.40 per 1M tokens$4.40Grok 4.7Grok 4.7: $6.00 per 1M tokens$6.00Claude Sonnet 5.5Claude Sonnet 5.5: $10.00 per 1M tokens$10.00GPT-6.1 SolGPT-6.1 Sol: $10.00 per 1M tokens$10.00

On the record

Model id
meta/llama-4-maverick-17b-128e-instruct
Context window
128K tokens
Max output
4K tokens
Input modalities
text, image
Output modalities
text
Reasoning
no
Tool calls
yes
Open weights
yes
Knowledge cutoff
Feb 2024
Released
2025-04-01
Last updated
2025-04-01

Back of the envelope

≈

← All Nvidia listings