Monday, October 5, 2026 · No. 435Prices synced 6 hours ago
All·AI·Model

The AI Model Price Index

Meta Llama 3.3 70B Versatile

by Helicone · family llama · listed Dec 2024T

Open Llama instruction model for multilingual chat, reasoning, and coding

Current rates USD per 1M tokens

Input$0.59
Output$0.79
Cache read—
Cache write—

Against the market

Input price vs. maker flagships ($/1M tokens)
DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.15 per 1M tokens$0.15Qwen3.8 Omni FlashQwen3.8 Omni Flash: $0.15 per 1M tokens$0.15Meta Llama 3.3 70B VersatileMeta Llama 3.3 70B Versatile: $0.59 per 1M tokens$0.59Gemini 3.8 FlashGemini 3.8 Flash: $0.75 per 1M tokens$0.75GLM-5.3GLM-5.3: $1.40 per 1M tokens$1.40Claude Sonnet 5.5Claude Sonnet 5.5: $2.00 per 1M tokens$2.00GPT-6.1 SolGPT-6.1 Sol: $2.00 per 1M tokens$2.00Grok 4.7Grok 4.7: $2.00 per 1M tokens$2.00
Output price vs. maker flagships ($/1M tokens)
Qwen3.8 Omni FlashQwen3.8 Omni Flash: $0.47 per 1M tokens$0.47DeepSeek V4 Flash Vision ExpDeepSeek V4 Flash Vision Exp: $0.60 per 1M tokens$0.60Meta Llama 3.3 70B VersatileMeta Llama 3.3 70B Versatile: $0.79 per 1M tokens$0.79Gemini 3.8 FlashGemini 3.8 Flash: $3.75 per 1M tokens$3.75GLM-5.3GLM-5.3: $4.40 per 1M tokens$4.40Grok 4.7Grok 4.7: $6.00 per 1M tokens$6.00Claude Sonnet 5.5Claude Sonnet 5.5: $10.00 per 1M tokens$10.00GPT-6.1 SolGPT-6.1 Sol: $10.00 per 1M tokens$10.00

On the record

Model id
llama-3.3-70b-versatile
Context window
131K tokens
Max output
33K tokens
Input modalities
text
Output modalities
text
Reasoning
no
Tool calls
yes
Open weights
no
Knowledge cutoff
Dec 2024
Released
2024-12-06
Last updated
2024-12-06
Provider docs
helicone.ai/models

Back of the envelope

≈

Also listed at

  • Groq$0.59 in · $0.79 out
  • Abacus$0.59 in · $0.79 out

← All Helicone listings