Friday, August 21, 2026 · No. 390Prices synced 15 hours ago
All·AI·Model

The AI Model Price Index

Llama 3.1 Nemotron Nano 8B v1

by Nvidia · family nemotron · listed Mar 2025RTW

Nemotron model for efficient reasoning, coding, and specialized AI agents

Current rates USD per 1M tokens

Inputfree
Outputfree
Cache read
Cache write

Against the market

Input price vs. maker flagships ($/1M tokens)
Llama 3.1 Nemotron Nano 8B v1Llama 3.1 Nemotron Nano 8B v1: free per 1M tokensfreeDeepSeek V4 ProDeepSeek V4 Pro: $0.435 per 1M tokens$0.435Gemini 3.7 FlashGemini 3.7 Flash: $0.75 per 1M tokens$0.75Mistral Medium 3.5Mistral Medium 3.5: $1.50 per 1M tokens$1.50Grok 4.6Grok 4.6: $2.00 per 1M tokens$2.00Qwen3.8 MaxQwen3.8 Max: $2.00 per 1M tokens$2.00Claude Opus 5Claude Opus 5: $5.00 per 1M tokens$5.00GPT-5.6 SolGPT-5.6 Sol: $5.00 per 1M tokens$5.00
Output price vs. maker flagships ($/1M tokens)
Llama 3.1 Nemotron Nano 8B v1Llama 3.1 Nemotron Nano 8B v1: free per 1M tokensfreeDeepSeek V4 ProDeepSeek V4 Pro: $0.87 per 1M tokens$0.87Gemini 3.7 FlashGemini 3.7 Flash: $3.75 per 1M tokens$3.75Grok 4.6Grok 4.6: $6.00 per 1M tokens$6.00Qwen3.8 MaxQwen3.8 Max: $6.00 per 1M tokens$6.00Mistral Medium 3.5Mistral Medium 3.5: $7.50 per 1M tokens$7.50Claude Opus 5Claude Opus 5: $25.00 per 1M tokens$25.00GPT-5.6 SolGPT-5.6 Sol: $30.00 per 1M tokens$30.00

On the record

Model id
nvidia/llama-3.1-nemotron-nano-8b-v1
Context window
131K tokens
Max output
16K tokens
Input modalities
text
Output modalities
text
Reasoning
yes
Tool calls
yes
Open weights
yes
Knowledge cutoff
Released
2025-03-18
Last updated
2025-03-18

Back of the envelope

← All Nvidia listings